TPU 8i Adds 288GB HBM and 384MB SRAM for Agent Inference
Google's eighth-gen TPUs split into TPU 8t for training (121 ExaFlops, 9,600 chips) and TPU 8i for inference (288GB HBM, 80% better perf-per-dollar).
5 posts
Google's eighth-gen TPUs split into TPU 8t for training (121 ExaFlops, 9,600 chips) and TPU 8i for inference (288GB HBM, 80% better perf-per-dollar).
Google's Gemini 3.8 Flash delivers frontier-level coding and reasoning at Flash speed and cost ($0.75/1M in tokens). A Cyber variant targets vulnerability discovery and patching.
Google's $920M/month SpaceX contract shifts AI infrastructure from on-demand cloud to locked-in capacity. Here's what changes for enterprise architecture.
Google released Gemma 4 on April 2, 2026 — four variants from 2B to 31B, with 256K context, native vision and audio, and Apache 2.0 licensing. Here's what it's for, where it fits, and how to run it.
Google's Gemini 3 Pro brings generative interfaces, 1M token context, and state-of-the-art multimodal reasoning to developers and consumers alike.