惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

酷 壳 – CoolShell
酷 壳 – CoolShell
aimingoo的专栏
aimingoo的专栏
P
Proofpoint News Feed
宝玉的分享
宝玉的分享
MyScale Blog
MyScale Blog
The GitHub Blog
The GitHub Blog
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
月光博客
月光博客
量子位
博客园 - 司徒正美
V
V2EX
I
InfoQ
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Vercel News
Vercel News
H
Hackread – Cybersecurity News, Data Breaches, AI and More
美团技术团队
N
Netflix TechBlog - Medium
L
LangChain Blog
IT之家
IT之家
Blog — PlanetScale
Blog — PlanetScale
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Stack Overflow Blog
Stack Overflow Blog
A
About on SuperTechFans
Microsoft Azure Blog
Microsoft Azure Blog

Interesting Engineering

US firm to scale laser-based nuclear fusion ‘breakthrough’ with new partnership Military Archives - Interesting Engineering World’s first non-nuclear lead-cooled reactor to generate electricity begins installation US scientists devise new process to turn sewage sludge into 99% pure natural gas US firm unveils submarine-hunting drone with 9,200-mile-range, 35 mph top speed Military Archives - Interesting Engineering Supercomputer finds lithium-titanium tweak to boost sodium-ion batteries for grids Lockheed Martin demonstrates vertical launch missile system for mobile drone defense China’s 1116 MWe Taipingling Unit 1 reactor goes online, set to generate 9bn kWh yearly ChatGPT Images 2.0 update combines reasoning, research, and design with 2K output US Navy tests plug-and-play laser system on USS Bush carrier, downs drones at sea China’s CATL reveals 621-mile EV battery, under-7-minute charging to challenge BYD US uses world’s first exascale supercomputer to model supernovae, fusion reactors AI and Robotics Archives - Interesting Engineering First-in-human study confirms safety of graphene-based brain interface Tesla’s Optimus humanoid robot greets runners, poses for photos at Boston Marathon Interlocking materials offer high strength and flexibility for robotics, infrastructure US redeploys 100,000-ton nuclear-powered aircraft carrier in Red Sea after repairs US scientists unveil concept for ‘world’s first neutrino laser’ to unlock breakthroughs New military tech can maintain communication in contested electronic warfare environments Got a dark personality? Psychologists can help you choose your career wisely Humidity boosts performance of 3D-printed nanogenerator instead of degrading it China demonstrates microwave beam that recharges drones in flight, continues power delivery Scientists run compact free-electron laser for eight hours, cracks FEL stability problem China’s PLA considers to use minelaying underwater drones to enforce Taiwan blockade: Report 1-ton sharks may struggle for survival in waters exceeding 62.6°F, study suggests US firm’s thorium nuclear fuel bundles move to manufacturing for commercial reactors Tesla hits 0% charge in remote Chilean desert as YouTuber uses hood-mounted solar Humanoid robot surpasses human world record in Beijing half-marathon, clocking 50:26 mins New method extracts maximum work from unknown quantum states using symmetry tricks
Google launches TPU 8 chips with 3x power to speed AI tra...
Neetika Walt · 2026-04-23 · via Interesting Engineering

Google has unveiled its eighth-generation Tensor Processing Units, introducing two custom AI chips designed separately for model training and inference as demand for large-scale AI computing surges.

Announced at Google Cloud Next, the new processors are called TPU 8t and TPU 8i. They are built to power Google’s AI Hypercomputer platform and support workloads ranging from training frontier models to serving AI agents in production.

TPUs are Google’s in-house accelerators that have powered internal systems such as Gemini for years. The company is now expanding that hardware to customers looking for alternatives to Nvidia-dominated AI infrastructure.

Google said both chips will become generally available later this year.

Two chips emerge

The TPU 8t is optimized for training large AI models. Google said a single superpod can scale to 9,600 chips and deliver 121 exaflops of compute performance.

The company added that TPU 8t offers nearly three times the compute performance per pod compared with the previous generation, Ironwood.

Training systems also received faster storage access and upgraded networking aimed at keeping chips busy instead of waiting for data.

Google said TPU 8t targets more than 97 percent “goodput,” a term used to measure productive compute time instead of idle time caused by failures or bottlenecks.

That matters because delays across massive clusters can add days to training schedules for advanced AI systems.

The TPU 8i focuses on inference, the stage where trained AI models answer prompts, run tools, and power software agents.

Agent era push

Google said TPU 8i includes 288 GB of high-bandwidth memory and 384 MB of on-chip SRAM, helping keep active model data closer to the processor for faster responses.

The chip also uses Google’s Axion Arm-based CPUs and upgraded interconnect bandwidth for Mixture of Experts, or MoE, models. These architectures activate only parts of a model at a time to lower costs while scaling performance.

According to Google, TPU 8i delivers 80% better performance-per-dollar than the prior generation, allowing customers to handle nearly twice the workload at the same cost.

The launch highlights how AI infrastructure is shifting beyond general-purpose GPUs toward specialized chips tuned for different workloads.

Google said the two-chip strategy was shaped by the rise of AI agents, which need systems that can reason through tasks, run workflows, and repeatedly interact with tools and other models.

In data centers, Google said both chips also offer up to two times better performance-per-watt than Ironwood.

They use fourth-generation liquid cooling to support higher compute density while controlling power use.

The announcement also underscores Google’s broader effort to challenge Nvidia’s grip on AI hardware by combining custom silicon, networking, software frameworks, and cloud services into one stack.

Both TPU 8t and TPU 8i will be available through Google Cloud later this year.

Google said the chips also support frameworks including JAX, PyTorch, SGLang, and vLLM, allowing developers to run existing AI workloads without major software rewrites or migration hurdles.

The Blueprint

Get the latest in engineering, tech, space & science - delivered daily to your inbox.

With over a decade-long career in journalism, Neetika Walter has worked with The Economic Times, ANI, and Hindustan Times, covering politics, business, technology, and the clean energy sector. Passionate about contemporary culture, books, poetry, and storytelling, she brings depth and insight to her writing. When she isn’t chasing stories, she’s likely lost in a book or enjoying the company of her dogs.