惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
T
The Blog of Author Tim Ferriss
博客园 - 司徒正美
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
有赞技术团队
有赞技术团队
量子位
S
SegmentFault 最新的问题
博客园 - 聂微东
博客园 - 【当耐特】
J
Java Code Geeks
美团技术团队
Hugging Face - Blog
Hugging Face - Blog
H
Help Net Security
V
V2EX
人人都是产品经理
人人都是产品经理
博客园 - Franky
罗磊的独立博客
Engineering at Meta
Engineering at Meta
A
About on SuperTechFans
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
酷 壳 – CoolShell
酷 壳 – CoolShell
云风的 BLOG
云风的 BLOG
Y
Y Combinator Blog
Apple Machine Learning Research
Apple Machine Learning Research

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Meta Burned 60 Trillion Tokens in 30 Days. Here Is How to...
Patrick Hugh · 2026-04-30 · via DEV Community

Meta built an internal leaderboard called "Claudeonomics." It tracked AI token consumption across 85,000 employees. Gamified tiers from bronze to emerald. Titles like "Token Legend" and "Session Immortal." A competitive race to use the most AI.

In 30 days, they burned 60 trillion tokens.

Then they shut it down.

What happened

The Claudeonomics dashboard was a voluntary internal tool on Meta's intranet. It ranked the top 250 AI token consumers with gamified incentives. The idea was to encourage AI adoption across the company.

It worked too well.

Multiple sources confirmed that employees left AI agents running for hours executing busywork research tasks specifically to climb the leaderboard. They consumed tokens while producing nothing of value.

The top individual consumer averaged 281 billion tokens per day. For a month straight.

Why it matters

Token consumption is an input metric. Not an output metric. Measuring productivity by tokens consumed is like measuring engineering quality by lines of code written.

Meta learned this the expensive way. But the lesson applies to every team running AI agents in production.

Here is the pattern:

  1. Team deploys AI agents
  2. No budget limits set
  3. Agents run autonomously (or employees run them to look productive)
  4. Token costs compound without anyone watching
  5. Someone notices a $50,000 cloud bill

Meta can absorb the cost. Your team probably cannot.

The math at your scale

Let's scale it down. Say you have 5 agents running production tasks. Each processes 100 requests per day. Average cost per request: $0.10.

That is $50/day. $1,500/month. Manageable.

Now one agent hits a retry loop. It fires 10,000 requests in an afternoon. That is $1,000 in one burst. No warning. No cap. Just a bill.

Or an agent starts looping through a research task with no termination condition. It runs all weekend. Monday morning, you have a $3,000 bill and a 2MB log file of circular reasoning.

This is not hypothetical. This is the default behavior of every agent framework that ships without budget controls.

What Meta should have done

Three things:

1. Budget limits per agent, per session

Every agent needs a hard cap. Not a soft warning. A hard stop.

from agentguard47 import init, BudgetGuard

init(
    guards=[BudgetGuard(max_cost=10.00)]
)

Enter fullscreen mode Exit fullscreen mode

When the budget hits $10, the agent stops. No negotiation. No override. The guard is deterministic. The agent cannot convince it to keep going.

2. Loop detection

Agents loop. It is what they do when they get stuck. Without detection, a loop runs until something external kills it (usually the credit card limit).

from agentguard47 import init, LoopGuard

init(
    guards=[LoopGuard(max_iterations=100)]
)

Enter fullscreen mode Exit fullscreen mode

100 iterations and done. If the agent has not solved the problem in 100 tries, iteration 101 is not going to help.

3. Kill switches

Sometimes you need to stop everything. Right now. Not "after the current batch finishes." Now.

AgentGuard's timeout guard gives you that:

from agentguard47 import init, TimeoutGuard

init(
    guards=[TimeoutGuard(max_seconds=300)]
)

Enter fullscreen mode Exit fullscreen mode

Five minutes. Then it is over. Combine all three for defense in depth.

The real lesson

Meta's Claudeonomics experiment failed because they measured the wrong thing. But the deeper failure was structural: 85,000 people running AI agents with no runtime budget controls.

The gamification just made the problem visible faster.

Every team running AI agents without budget limits is running the same experiment. You just do not have a leaderboard showing you the results.

Set your limits before you need them. Not after.


AgentGuard is an open-source Python SDK for AI agent runtime safety. Budget limits, loop detection, and kill switches. Zero dependencies. Local-first.

Get started with AgentGuard

Related: AI Agent Cost and Pricing in 2026