惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

人人都是产品经理
人人都是产品经理
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
宝玉的分享
宝玉的分享
月光博客
月光博客
爱范儿
爱范儿
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
WordPress大学
WordPress大学
有赞技术团队
有赞技术团队
阮一峰的网络日志
阮一峰的网络日志
博客园_首页
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
博客园 - 三生石上(FineUI控件)
博客园 - 聂微东
小众软件
小众软件
量子位
MongoDB | Blog
MongoDB | Blog
Blog — PlanetScale
Blog — PlanetScale
The Cloudflare Blog
Stack Overflow Blog
Stack Overflow Blog
U
Unit 42
Hugging Face - Blog
Hugging Face - Blog
T
The Blog of Author Tim Ferriss
H
Help Net Security
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC

Amazon News

How Amazon is making its data centers more water-efficient 20 best books of 2026 so far, according to Amazon editors Claude Fable 5 from Anthropic now available on Amazon Bedrock Amazon's impact in Seattle, Bellevue, and the Puget Sound: Amazon employees volunteer for spring cleanups across the Puget Sound Amazon in the community: Service, community, and commitment at HQ2 Amazon Leo's largest Arianespace launch yet will feature upgraded boosters to deploy more satellites 'Vought Rising': How to watch on Prime Video Amazon disaster relief technology: 2,000 free systems that restore power, Wi-Fi, and water in minutes Pinterest signs $4 billion AWS deal to scale AI for 600 million users Amazon announces new robots, faster delivery, and 25,000 jobs in Europe Amazon unveils new Proteus robot in €10 billion European investment in its fulfillment centers Amazon customers in the UK can now add fresh groceries to Same-Day Delivery orders in parts of London Prime Video's most-watched original films and series last week worldwide NBA Finals 2026: How to watch Knicks vs. Spurs live on Prime Video What does an Amazon fulfillment center GM do every day? Amazon launches Prime in South Africa Spider-Man: Brand New Day early screenings in theaters through Prime How AWS transformed IT by making cloud computing accessible to everyone The best early Prime Day 2026 deals to shop now Amazon Prime Day 2026: Ofertas exclusivas para miembros Prime del 23-26 de junio When is Amazon Prime Day 2026? Shop deals June 23-26 What’s new on Prime Video in June 2026, including ‘The Legend of Vox Machina,’ NWSL Challenge Cup, and more Sustainability at Amazon: Water conservation efforts, supports California watershed restoration through innovative Forest Resilience Bond How Amazon fights freight fraud and protects products How Amazon is investing in Indiana's economy What is agentic AI? Amazon's approach to reliable AI agents How AWS used random graph theory to build more efficient data centers Amazon Japan launches Shinkansen bullet train deliveries Amazon Health Services welcomes Dr. Roy Schoenberg as new SVP AWS Agentic Shopping Assistant: Amazon's AI shopping tech, now for any retailer
How Amazon CPUs like AWS Graviton enable agentic AI
2026-04-24 · via Amazon News

Why CPUs matter for agentic AI

Agentic AI systems reason continuously and make real-time decisions—a fundamental shift that requires different infrastructure than traditional AI training.

An AWS Graviton4 chip.

Amazon Staff

Written by Amazon Staff

3 min read

Key takeaways

  • Agentic AI operates continuously, not in batches, requiring sustained computing power with fast communication between processing cores.
  • CPUs like AWS Graviton are designed for these always-on reasoning workloads.
  • Companies like Meta are deploying tens of millions of Graviton cores to power agentic AI at global scale.

What makes agentic AI different

For years, conversations about AI infrastructure centered on chips designed for training large language models. Accelerators like AWS Trainium and graphics processing units (GPUs) excel at processing massive amounts of data in parallel, making them ideal for the intensive work of teaching AI models.

But agentic AI operates differently. To understand why the central processing unit (CPU) is reclaiming relevance in the AI conversation, we have to look at the fundamental difference between a standard language model (SLM), large language model (LLM) and an AI agent. A language model is almost like a calculator: you give it a prompt, and it performs a massive amount of parallel math to predict the next set of tokens (the output). This is where GPUs shine—they run thousands of calculations simultaneously across many cores (individual processing units within a chip).

An AI agent, however, is more like a manager. Agents are AI systems capable of acting autonomously to complete multistep tasks—always on and continuously processing information and taking action. Think of a digital assistant that doesn't just respond to commands but actually completes tasks on your behalf—managing schedules, coordinating across systems, making decisions. If you ask an agent to “research this company and draft a brief,” it doesn’t just generate text. It must break the goal into sequential steps, open a web browser and navigate links, parse files, data, and filter noise, and run code to finalize the draft.

x

While inference has long been the GPU’s domain, (with CPUs now handling a growing share of that workload too), every other step in that chain, from logic to file management to network calls and code execution, is a CPU-native task. That requires sustained computing power with extremely fast communication between processing cores.

This is where purpose-built CPUs like AWS Graviton become critical. Graviton processors are designed for these continuous, low-latency workloads that define the agentic AI era. That's why companies like Meta are using Graviton for their latest agentic AI workloads—deploying tens of millions of Graviton cores to power AI systems that reason continuously and serve billions of users.

How Graviton addresses the challenge

Agentic systems engage in rapid execution cycles—not just reasoning, but actively retrieving data, calling tools, and taking action before looping back to evaluate next steps. Each cycle requires different parts of the processor to share data quickly.

Graviton chips are designed to minimize the time different parts spend communicating with each other. For AI systems constantly exchanging information during reasoning processes, that speed matters.

Organizations running agentic AI at global scale are choosing Graviton because the architecture matches the workload. Training a model happens in intense bursts. Running an agentic system requires sustained performance and constant adaptation—different work needs different tools.

Why this matters for continuous intelligence

When AI systems serve millions of users continuously, performance and energy efficiency both become essential. Graviton delivers better performance and is our most energy efficient processor. For systems running around the clock at global scale, that efficiency is both environmentally conscious and economically essential. The combination makes it viable to run sophisticated agentic systems continuously for large user bases.

A customer entering the back of a car with an AWS and Uber logo overlay.

Building infrastructure for systems that never stop thinking

The rise of agentic AI represents a fundamental shift in AI infrastructure. As AI systems become more autonomous and more integrated into daily digital experiences, the requirements evolve.

Processors like Graviton—designed for sustained, low-latency computing rather than periodic training bursts—are becoming the foundation for this new era. The question isn't whether agentic AI will reshape how we work and live. It's whether the infrastructure exists to support it at scale.

Increasingly, that infrastructure is being built on technology designed specifically for systems that never stop thinking.

Learn more about how Graviton chips make cloud computing faster, cheaper, and more energy efficient.

More Amazon News

1 / 1