惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

IT之家
IT之家
Last Week in AI
Last Week in AI
博客园_首页
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园 - 叶小钗
大猫的无限游戏
大猫的无限游戏
人人都是产品经理
人人都是产品经理
V
Visual Studio Blog
宝玉的分享
宝玉的分享
博客园 - Franky
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
月光博客
月光博客
T
Tailwind CSS Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
量子位
博客园 - 聂微东
S
SegmentFault 最新的问题
博客园 - 司徒正美
罗磊的独立博客
V
V2EX
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
美团技术团队
小众软件
小众软件
Jina AI
Jina AI

Show HN

GitHub - astefanutti/shaderbang: Shebang for Shaders Show HN: Generate Claude Code Workflows using Spec Driven Development approach Show HN: AI agents for UK GDAD PCF roles and their skills The Two Pillars: Mixer Mode and Meta-Software in the Reorganization of Software Work After AI GitHub - JaiCode08/teleport-env What 1,000+ Harness Experiments Taught Me About Self-Improving Agents Show HN: Liiists, a Markdown-first, iOS and CLI list app SwiperTab – Get this Extension for 🦊 Firefox (en-US) GitHub - kouhxp/fftext: Summarize, explain, fact-check, or translate any text, URL, or file. No GPU. No cloud. One command GitHub - sweetpad-dev/sweetpad: Develop Swift/iOS projects using VSCode GitHub - dogmaticdev/IRON: IRON a.k.a. Intermediate Representation Object Notation is a Interpreter/Database that is used to create Programming Languages. GitHub - sjhalani7/vaen: Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config, or secrets. Show HN: Gandalf the Grader Show HN: Citadeld – replay any CI failure locally from a single file GitHub - tdortman/cuSBF: High-Performance GPU Super Bloom Filter coral-ai/claude-code-token-xray at main · Coral-Bricks-AI/coral-ai GitHub - ulyssestenn/funes: Funes is a Git-based framework for LLM-managed knowledge work: an AI Librarian ingests raw sources, builds an interlinked Markdown knowledge base, and uses it to produce cited reports, analyses, and other outputs. GitHub - ThatXliner/gah: Git Add Hunk, built for agents to use GitHub - harmont-dev/harmont-cli: Command-line client for the Harmont CI platform GitHub - brooksmcmillin/mcp-authflow: OAuth 2.0 Authorization Server framework for MCP servers GitHub - javaid-codes/audit-supply-chain-agents GitHub - amorey/gochan: A small library of common channel architectures for Go, inspired by Rust GitHub - arifozgun/OpenGem: Free, Open-Source AI API Gateway with Gemini, OpenAI & Anthropic Compatibility in 1 file GitHub - Pranesh950/BioPetals: 🌸 Run BIOxAI models at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading GitHub - cnguyen14/bounty-doctor: Diagnose a GitHub bounty issue before you waste hours: detects honeypot scam repos, AI-bot attempt swarms, and stale contests. Show HN: CoreMCP – MCP Server for On-Prem DBs Show HN: KittyHTML – Render HTML/CSS as an inline image in your terminal GitHub - bingud/filemat: Web-based file manager Show HN: TruthLens – Free multi-signal deepfake image detector GitHub - apexlocal-jz/claude-usage-tray: Windows system-tray app showing your Claude Code rate-limit usage at a glance. Zero deps, ~300 lines of PowerShell. Cross-IDE (works regardless of VS Code, Cursor, plain terminal).
Cut Claude Code Costs ~50% Without Quality Loss | Headroom
gghootch · 2026-05-29 · via Show HN

Cut your Claude Code token costs by ~50%

Headroom is a menu bar app that quietly optimizes the inputs Claude Code gets by trimming prompt bloat, stripping boilerplate, and compressing documents without changing how you work.
This unlocks about 2x as much Claude Code usage on the Claude plan you already pay for.

181 developers using Headroom

11.8B tokens saved, and counting

~$105,997 saved at Claude API rates

Privacy first

Your prompts never touch our servers — everything runs locally on your machine.

Self-contained

Keeps your runtime clean, never interfering with packages your projects depend on.

Fewer tokens, same result

Smart optimization cuts noise before Claude Code sees it, with no impact on the output.

How it works

Less noise in, more Claude out

Headroom intercepts every prompt before it reaches Claude, strips out logs, boilerplate, and repetitive content, then forwards only what the model needs — cutting your token spend by ~50% without impacting output quality.

Your tools

logsHTMLJSONshell

Claude Code

sees only what matters

65.1M

tokens saved per developer, on average

Headroom savings dashboard

Benchmarks

Same results, fewer tokens

Headroom compresses aggressively — but without throwing anything away. Real workloads, real results, measured before and after.

Token savings by scenario

Build log (200 lines) 93.9% saved

148 remaining 2,264 tokens saved

JSON array (100 items) 90.6% saved

297 remaining 2,866 tokens saved

Shell output (200 lines) 85.5% saved

469 remaining 2,769 tokens saved

JSON array (500 items) 83.1% saved

1,614 remaining 7,912 tokens saved

Multi-tool agent (memory leak investigation) 61% saved

6,100 remaining 9,562 tokens saved

Headroom powered savings

Tokens sent after optimization

Quality preserved

Fewer tokens doesn't mean fewer answers. Headroom strips noise — not signal. Every benchmark below ran the same task with and without compression, then compared the outputs.

0.919

HTML extraction F1

181 real web pages (Scrapinghub)

4/4

JSON retrieval

needle-in-haystack, 100 prod logs

+0.02 F1

QA accuracy vs. uncompressed baseline

Stripping HTML noise helped the model focus on relevant content — compression improved results on SQuAD v2 / HotpotQA (+2% exact match).

0%

HTML recall

181 real web pages (Scrapinghub)

Same

Multi-tool agent findings

4-tool session, memory leak task — identical conclusions at 61% fewer tokens

Based on data from the open-source Headroom CLI benchmark suite.

ROI Calculator

See what Headroom saves your team

Headroom costs a fraction of your Claude subscription and delivers roughly twice the usage.

Claude Pro Max ×5 Max ×20

Engineers using Claude Code 10

151025501002505001000

$1,000

Monthly Claude spend

Equivalent extra capacity

$1,000/mo

8× return on Headroom spend — based on ~2× token efficiency from Headroom.

Testimonials

Loved by developers

Pricing

Plans for every Claude tier

Create a Headroom account to unlock your 14-day trial, then choose the plan that matches your Claude tier. Need rollout controls or private deployment? Talk to us about Headroom for teams.

Includes:

  • Unlock cost savings and stats
  • Up to 50% of your weekly limit
  • Optimize Claude Code practices

Everything in Free, plus:

  • Unlimited use with Claude Pro
  • Track sessions across devices
  • Email-based support

Includes:

  • Use with Claude Max x5
  • Track sessions across devices
  • Email-based support

Includes:

  • Use with Claude Max x20
  • Track sessions across devices
  • Priority support

Built on Headroom CLI

Headroom for desktop is built on Headroom CLI.

The Headroom desktop app is based on the open-source Headroom CLI project created by Tejas Chopra.
The desktop app is created with the endorsement and support of Tejas.

Resources

Learn how to lower Claude Code costs

Guides on reducing Claude Code costs, understanding usage limits, and cutting Claude API spend — plus a product FAQ for privacy, quality, and rollout questions.

Cost Guide

How to reduce Claude Code costs

Learn where token waste comes from, which workflows benefit most from compression, and how Headroom helps preserve quality while cutting spend.

Usage Guide

Claude Code usage: what counts and how to get more from your plan

Learn what burns usage fastest, what counts toward your plan, and how to make the same Claude tier last longer.

Why So Expensive

Why is Claude Code so expensive?

The four patterns that drive Claude Code token spend: verbose tool output, repeated context, multi-step debugging, and large codebase reads.

Usage Limits

Claude Code usage limits and the 5-hour window

How the 5-hour rolling window and weekly cap work, what each plan covers, and how to keep coding without immediately upgrading.

Claude API

Reduce Claude API costs in 2026

Practical levers for cutting Claude API spend — prompt caching, model tier routing, output limits, batch API — plus the Claude Code shortcut.

FAQ

Headroom FAQ for Claude Code savings

Get quick answers about local processing, supported platforms, benchmarks, and how to evaluate whether Headroom fits your team.

Ready to try it?

Start with Headroom for free

Install the app, connect your account, and start reclaiming Claude Code usage in minutes.