惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

雷峰网
雷峰网
Y
Y Combinator Blog
酷 壳 – CoolShell
酷 壳 – CoolShell
The Cloudflare Blog
博客园_首页
J
Java Code Geeks
A
About on SuperTechFans
人人都是产品经理
人人都是产品经理
量子位
C
Check Point Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
博客园 - 三生石上(FineUI控件)
L
LangChain Blog
N
Netflix TechBlog - Medium
Hugging Face - Blog
Hugging Face - Blog
B
Blog
美团技术团队
Microsoft Security Blog
Microsoft Security Blog
P
Proofpoint News Feed
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
宝玉的分享
宝玉的分享
罗磊的独立博客
MongoDB | Blog
MongoDB | Blog
Last Week in AI
Last Week in AI

Exponential View

🔮 What would Adam Smith make of AI? 📈 Anthropic’s $517 billion shopping list 🔮 Look up, the curve turned 📈 AI revenue hit $229 billion 🔮 Astra, the good, the bad and the ugly EV #600 📈 Data to start your week 🔮 The containment era #599 📈 Data to start your week 🏦 The problem with petards 🫧 Is AI a bubble yet? Our five gauges say no 🔮 Introducing: AI Economy Research Fellowship 📈 Data to start your week 🔮 The curious economics of a $6 AI agent #597 What the Google DeepMind exodus tells us about the AI cycle 📈 Making sense of the AI capex logjam 🔮 Agents form alliances, DeepMind’s reset & how likely is a crash? #596 🔮 Seven lessons for managing AI agents 📈 Data to start your week 🔮 Leopold & exponential markets; transformative GLP-1s; runaway AI & the future of safety++ 📚 My non-obvious summer reading list 🔮 For AI adopters, success and failure looks the same right now 📈 Data to start your week 🔮 The curious case of AI distillation 🔮 Will Kimi K3 change the economics of AI? 📈 Data to start your week 🔮 Kimi’s positive impact. Why are solar costs going up? AI & copyright ++ #593 📈 Data to start your week 🔮 AI & the great unglobalization 📈 Data to start your week 🔮 Exponential View #591: Never skilling; China’s self-reliance; screwworm & progress; synth cells, tungsten & cheating AI++
🔮 Why one AI is better than four #598
Azeem Azhar · 2026-08-23 · via Exponential View

Good morning!

We are looking for an outstanding economist to join us as an AI Economy Research Fellow. If you know someone we should speak to, send them our way.

A few months ago, we (alongside Rohit Krishnan) looked at whether AI is immune to groupthink. The answer was no. Blending several models’ answers kept about a quarter of the good ideas that had come from a single model. This is called the hidden-profile problem: when groups discuss what everyone already knows and don’t get to the knowledge that only one member holds. Anthropic has now run that classic experiment on agents: four agents must arrive at a decision. The evidence they hold in common points to the wrong option, while only a few agents (or just one) have the facts that lead to a correct decision. Getting it right means a small set of agents pressing its private facts and the others trusting them over the apparent consensus. After discussion, most model families chose correctly in only 17-36% of runs, while a single agent handed the entire evidence base got it right nearly every time. Only one model (somewhat) escaped: Mythos 5, at about 85% (why, we don’t know).

I see two problems at work here. First, LLMs lack diversity (they are low-variance): set 30 agents the same coding task and 18 of them will name their git branch identically. Second, agents lack the institutions that make human groups robust: reputation, recourse and protection for the lone dissenter. These aren’t necessarily unfixable, but it’s not yet clear what the fix is. On the diversity side, I particularly like the solutions Thinking Machines puts forward: an ecosystem of AIs raised in different places, with different values and purposes, “keeping the weirdness alive.” After all, most good ideas started weird.

In our State of AI report, we found a positive but underwhelming elasticity for tokens. A 10% price cut lifts token use by 12–18%: enough to raise total spend, but not by much.

Patrick Saner made a comment that made me rethink why: “the cost per token is irrelevant. What matters is the cost of completing a useful unit of work.” Elasticity might be underwhelming because users haven’t found a way to properly price “a useful unit of work.” Firms exist exactly to avoid pricing work. Especially for knowledge work, we buy a lot of it in bundles: a salary, a retainer, an hour. Creating a priceable task from knowledge work is not easy. Some may have found a useful unit: since October 2023 the top 1% of firms raised AI spend per employee by $6,542. The median rose only $9.63. I would guess this is mostly software, where AI is both most proven and, in a sense, most measurable (commits, pull requests and releases).