惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

人人都是产品经理
人人都是产品经理
有赞技术团队
有赞技术团队
WordPress大学
WordPress大学
月光博客
月光博客
T
Tailwind CSS Blog
阮一峰的网络日志
阮一峰的网络日志
小众软件
小众软件
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Last Week in AI
Last Week in AI
大猫的无限游戏
大猫的无限游戏
S
SegmentFault 最新的问题
罗磊的独立博客
Jina AI
Jina AI
酷 壳 – CoolShell
酷 壳 – CoolShell
宝玉的分享
宝玉的分享
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 三生石上(FineUI控件)
量子位
雷峰网
雷峰网
Apple Machine Learning Research
Apple Machine Learning Research
美团技术团队
博客园 - 聂微东
V
V2EX

Hacker News: Front Page

SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Introducing Claude Opus 4.7 Qwen Studio The Future of Everything is Lies, I Guess: Where Do We Go From Here? GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis Ancient DNA reveals pervasive directional selection across West Eurasia [pdf] AI cybersecurity is not proof of work Moving a large-scale metrics pipeline from StatsD to OpenTelemetry / Prometheus GitHub - Nightmare-Eclipse/RedSun: The Red Sun vulnerability repository GitHub - SethPyle376/hiraeth: Local AWS emulator focused on fast integration testing, with SQS support, SQLite-backed state, and a debug-friendly web UI. A Better Ludum Dare; Or, How to Ruin a Legacy GitHub - macOS26/Agent: Any AI, replaces Claude Code, Cursor, OpenClaw. Over 18 LLM providers (Claude, OpenAI, Gemini, Ollama, Zai, HF, Qwen) wired into a native Mac app that writes code, builds Xcode projects, bumps versions, manages git, automates Safari, use AppleScript, JS or Accessibility, extend Agent! w/ MCP Servers, run tasks from your iPhone via Messages. YouTube now lets you turn off Shorts I Made a Terminal Pager Burgers | マクドナルド公式 Commands — HackerNews CLI documentation ChatGPT for Excel PiCore - Raspberry Pi Port of Tiny Core Linux Live Nation illegally monopolized ticketing market, jury finds Google Broke Its Promise to Me. Now ICE Has My Data. Founding Engineer at Adaptional | Y Combinator CRISPR takes important step toward silencing Down syndrome’s extra chromosome GitHub - saffron-health/libretto: The AI toolkit for building reliable browser automations US v. Heppner (S.D.N.Y. 2026) no attorney-client privilege for AI chats [pdf] Unexpected €54k billing spike in 13 hours: Firebase browser key without API restrictions used for Gemini requests Fragments: April 14 Cal.com Goes Closed Source: Why AI Security Is Forcing Our Decision | Cal.com - Scheduling Software for Online Bookings Laravel raised money and now injects ads directly into your agent Codex Hacked a Samsung TV
How the hell is Groq raising more money?
zach · 2026-06-02 · via Hacker News: Front Page

Axios just dropped a bizarre scoop. Groq, the AI chip company company that was acquired by Nvidia in December of last year, is raising $650M. How, exactly, is a company that successfully exited raising more capital?

Well, technically, Nvidia did not acquire Groq. They licensed Groq’s technology and hired Groq’s key technical executives, but did not acquire the Groq corporate entity. That corporate entity continued to operate, focusing on maintaining Groq’s datacenters and their inference API. That API focuses on offering extremely fast inference on smaller models; their largest supported model is GPT OSS 120B, which is likely at least 10x smaller than frontier models like GPT-5.5 or Claude Mythos. This is a technical limitation of Groq’s architecture; without large amounts of high-bandwidth memory (HBM) in each chip package, the total cost to build and maintain a Groq cluster capable of serving a frontier model would be massive. However, their all-SRAM strategy lets them serve small models much faster than a conventional HBM-based chip, delivering more tokens-per-second at the cost of lower tokens-per-dollar. For certain applications, this makes sense.

More importantly, Groq has four large datacenter deployments that are already set up to serve inference workloads at scale. Building out new datacenters has been a major challenge for startups and hyperscalars alike, so Groq’s datacenters represent a major strategic asset. However, it remains to be seen if the Groq name gives them a legitimate leg up, and or if their brand association with the LPU chip and a specific high-speed, high-cost inference strategy limits them.

As inference demand surges, existing datacenters are quickly reaching full utilization and companies are trying to build new ones as fast as possible. But between regulation, power issues, and expertise, even large companies with experience building out datacenters are experiencing major delays in their datacenter construction projects. And for venture investors, who primarily invest in startups, getting direct exposure to datacenter demand is very difficult; because datacenters are so hard to build, very few startups are actually doing so.

Groq already has four functional datacenters, as well as the talent to build and operate more. When Nvidia licensed Groq’s technology, Groq’s chip design team, compiler team, and software team joined Nvidia, but Groq’s datacenter team stayed to maintain GroqCloud inference services. This makes the remaining Groq company a unique asset -- a private inference datacenter operator with clear operational expertise, plus what is likely an extremely low valuation due to most of their differentiating technology being acquired by Nvidia.

Let’s look at the numbers for some publicly traded AI datacenter companies to compare. CoreWeave is a 50 billion dollar company operating 43 AI datacenters. Nebius is worth about 50 billion dollars too, with only 11 datacenters, though they are larger than CoreWeave’s. One could argue that Groq’s datacenters alone could make them worth billions of dollars.

Thanks for reading zach's tech blog! This post is public so feel free to share it.

Share

Obviously, there are a couple issues with comparing Groq to CoreWeave and Nebius directly. Groq’s datacenters are all full of LPUv1 chips, which are seven years old at this point. And the new LPUv3 chips based on Groq’s architecture are being sold by Nvidia to any cloud provider who wants them. That means that Groq’s biggest technical advantage, of very fast tokens, is no longer unique to them.

Plus, the value of Groq’s brand adds additional complexity to the calculation. On one hand, Groq was a massively successful outcome. On the other, getting “acquired” and then still operating is confusing. And their brand is strongly associated with ultra-high-speed inference, which may not end up dominating inference workloads compared to lower-speed batched inference. This is particularly worrying as major organizations like Microsoft and Uber are raising the alarm about the high cost of AI coding tools.

Maybe Nvidia is giving Groq a sweetheart deal on buying hardware based on Groq’s technology, which would make Groq a significantly more appealing investment. Ultimately, it remains to be seen if Groq can successfully outfit their existing datacenters with new hardware to remain competitive with other cloud providers. I’m also curious to see whether they stay committed to high-speed, high-cost tokenomics, and if that strategy ends up being successful or not in the long term.

Discussion about this post

Ready for more?