惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

IT之家
IT之家
博客园 - 聂微东
雷峰网
雷峰网
Microsoft Azure Blog
Microsoft Azure Blog
WordPress大学
WordPress大学
Hugging Face - Blog
Hugging Face - Blog
S
SegmentFault 最新的问题
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
T
Tailwind CSS Blog
博客园 - 三生石上(FineUI控件)
V
Visual Studio Blog
博客园 - 司徒正美
爱范儿
爱范儿
月光博客
月光博客
阮一峰的网络日志
阮一峰的网络日志
博客园_首页
博客园 - 【当耐特】
Jina AI
Jina AI
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
酷 壳 – CoolShell
酷 壳 – CoolShell
大猫的无限游戏
大猫的无限游戏
小众软件
小众软件
人人都是产品经理
人人都是产品经理
V
V2EX

SiliconANGLE

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws - SiliconANGLE AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle - SiliconANGLE Team Cymru launches Total Insights Feed to replace legacy threat intelligence lists - SiliconANGLE AI Mode in Chrome adds split-screen view to enhance the web search experience - SiliconANGLE Resolve AI raises $40M at $1.5B valuation to optimize production environments - SiliconANGLE How Zscaler and OpenAI turn zero-trust security into an AI accelerator - SiliconANGLE OpenAI ratchets up Codex's agentic capabilities to rival Claude Code - SiliconANGLE Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements - SiliconANGLE Slash raises $100M at a $1.4B valuation to expand AI-powered banking platform for online businesses - SiliconANGLE Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work - SiliconANGLE Data center, consumer device chips boost TSMC’s revenue - SiliconANGLE Mission-critical security cannot be bolted on, says Oracle - SiliconANGLE Agentic infrastructure reshapes enterprise AI - SiliconANGLE Data quality, and data freedom, foundational for AI success - SiliconANGLE Data trust is a bedrock in successful, scalable AI outcomes - SiliconANGLE Google introduces new agentic AI-ready tools and resources for Android developers  - SiliconANGLE Agentic AI orchestration separates winners from laggards - SiliconANGLE Data-driven tools turning the tide against human trafficking - SiliconANGLE Achieving trusted AI development goes beyond 'vibes' - SiliconANGLE Impinj boosts edge computing power in updated R700 RAIN RFID reader - SiliconANGLE Certinia powers professional services with AI - SiliconANGLE Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M - SiliconANGLE Developer tooling startup Expo nabs $45M investment - SiliconANGLE Solidroad lands $25M to bring AI to customer support interactions - SiliconANGLE DuploCloud lands compliance and AI governance certifications as enterprise buyers tighten scrutiny - SiliconANGLE Lua lands $5.8M to help businesses build and manage AI agent workforces - SiliconANGLE Best of frenemies: Oracle's and AWS' clouds unite with dedicated, private connectivity - SiliconANGLE NIST shifts National Vulnerability Database to risk-based triage as CVE submissions hit record levels - SiliconANGLE Cisco goes to the races with new Churchill Downs multiyear partnership - SiliconANGLE Susecon 2026 will tackle the future of open-source platforms - SiliconANGLE
Inference chip startup Groq raises $650M to grow its clou...
by Maria Deutscher · 2026-06-23 · via SiliconANGLE

UPDATED 16:06 EDT / JUNE 22 2026

INFRA

Inference chip startup Groq raises $650M to grow its cloud platform

Seven months after inking a $20 billion chip licensing deal with Nvidia Corp., Groq Inc. today announced that it has raised $650 million in funding.

Growth investment firm Disruptive and hedge fund Infinitum led the round.

Groq has developed a chip design called the LPU that’s specifically optimized for artificial intelligence inference workloads. In December, Nvidia agreed to license the technologies that underpin the processor. It also hired several key Groq employees, including its founding chief executive.

The transaction produced the Nvidia Grok LPU 3, an inference processor that the chip giant debuted in March. It ships as part of a rack-size, liquid-cooled appliance called the LPQ. The system includes 32 trays that each host three Groq LPU 3 units, one central processing unit and network equipment.

The accelerators in an inference cluster each include a quartz crystal called a clock that regulates processing speeds. Clocks also play an important role in coordinating the flow of data between chips. When accelerators’ clocks move out of sync with each other, data traffic slows down, which negatively impacts AI model response times.

The LPU 3 includes a feature that automatically fixes clock drift to avoid data traffic bottlenecks. According to Nvidia, the chip includes 92 lanes that can each move data to other processors at a speed of 112 gigabits per second. That translates to 2.5 terabits per second of bidirectional bandwidth.

Accelerating the flow of data between chips is not the only way the LPU 3 speeds up inference workloads. The processor ships with 500 megabytes of onboard SRAM, a high-speed memory variety. SRAM is more performant than the off-chip RAM that other AI accelerators use to store data, which translates into faster inference.

Groq operates an LPU-powered cloud platform that companies can use to run inference workloads. The company disclosed today that the platform is processing trillions of tokens per week for 5 million developers.

Groq’s cloud runs across 13 data centers spanning multiple continents. The company will use the proceeds from its funding round to grow its inference capacity with the goal of reaching 200 megawatts by 2027. According to Groq, some of the new processing power will be provided by the LPX, the liquid-cooled LPU 3 appliance that Nvidia debuted in March.

Other cloud operators can theoretically build LPQ-powered inference services of their own. One way Groq could set itself apart from such potential rivals is by extending its platform with new services such as managed databases. Other AI-focused cloud providers, notably CoreWeave Holdings Inc., have also broadened their focus beyond infrastructure to higher-level services.

Image: Groq

A message from John Furrier, co-founder of SiliconANGLE:

Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities.

  • 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more
  • 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network.

About SiliconANGLE Media

SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement. As the parent company of SiliconANGLE, theCUBE Network, theCUBE Research, CUBE365, theCUBE AI and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.

Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.