惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

T
Troy Hunt's Blog
Blog — PlanetScale
Blog — PlanetScale
Engineering at Meta
Engineering at Meta
F
Full Disclosure
Recorded Future
Recorded Future
The GitHub Blog
The GitHub Blog
Microsoft Security Blog
Microsoft Security Blog
GbyAI
GbyAI
博客园_首页
博客园 - 叶小钗
MongoDB | Blog
MongoDB | Blog
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Recent Commits to openclaw:main
Recent Commits to openclaw:main
H
Hacker News: Front Page
人人都是产品经理
人人都是产品经理
The Cloudflare Blog
博客园 - 司徒正美
Webroot Blog
Webroot Blog
Google DeepMind News
Google DeepMind News
Help Net Security
Help Net Security
Cloudbric
Cloudbric
PCI Perspectives
PCI Perspectives
有赞技术团队
有赞技术团队
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
TaoSecurity Blog
TaoSecurity Blog
L
Lohrmann on Cybersecurity
量子位
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
T
Tailwind CSS Blog
Hacker News - Newest:
Hacker News - Newest: "LLM"
B
Blog RSS Feed
Apple Machine Learning Research
Apple Machine Learning Research
大猫的无限游戏
大猫的无限游戏
P
Proofpoint News Feed
N
News and Events Feed by Topic
罗磊的独立博客
T
Threat Research - Cisco Blogs
Schneier on Security
Schneier on Security
T
Tor Project blog
IT之家
IT之家
M
MIT News - Artificial intelligence
S
Security @ Cisco Blogs
O
OpenAI News
AI
AI
S
Securelist
Simon Willison's Weblog
Simon Willison's Weblog
The Last Watchdog
The Last Watchdog
月光博客
月光博客
Security Archives - TechRepublic
Security Archives - TechRepublic
L
LINUX DO - 热门话题

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant Common SOC 2 Failures (Real World) Stop Vibe-Checking Your AI App: A Practical Guide to Evals How to Use SonarQube and SonarScanner Locally to Level Up Your Code Quality Your Next To-Do App Is Dead — I Replaced Mine with an OpenClaw AI Sign a Nostr event in 60 lines of Python using coincurve — no nostr-sdk, no nbxplorer, no rust toolchain ITGC Audit Explained Like You’re in Big 4 Patch Tuesday abril 2026: Microsoft parcha 163 vulnerabilidades y un zero-day en SharePoint Stop scraping everything: a better way to track competitor price changes Listing on MCPize + the Official MCP Registry while routing payments OUTSIDE the marketplace — how I kept 100% of my x402 revenue Building an AI-Powered Risk Intelligence System Using Serverless Architecture Why We Ripped Function Overloading Out of Our AI Toolchain Testing AI-Generated Code: How to Actually Know If It Works SaaS Churn Is Killing Your Business. Here Is What to Do About It (Without a Support Team) The Speed of AI Is No Longer Linear - And Self-Improving Models Are Why How to Implement RBAC for MCP Tools: A Practical Guide for Engineering Teams From Standard Quote to Persuasive Proposal: AI Automation for Arborists I built a CLI that scaffolds complete multi-tenant SaaS apps Axios CVE-2025–62718: The Silent SSRF Bug That Could Be Hiding in Your Node.js App Right Now The dashboard that ended our friendship Data Pipelines Explained Simply (and How to Build Them with Python) The Hidden Cost of AI Systems Nobody Talks About. undefined vs undeclared, and how typeof behaves Switching from file-based jobs to NATS/Kafka in Rust without changing code io_uring Adventures: Rust Servers That Love Syscalls Why Agentic AI is Killing the Traditional Database The POUR principles of web accessibility for developers and designers Quantum Neural Network 3D — A Deep Dive into Interactive WebGL Visualization How To Install Caveman In Codex On macOS And Windows Automation Pipeline Reliability: Why Your Workflow Breaks When Nobody Is Watching I Built an 'Open World' AI Coding Agent — It Works From ANY Folder From Freelancing to Product: A Tech Service Company's SaaS Transformation China's AI Giants: Adding Tencent Hunyuan & ByteDance Doubao to AI University (74 Providers) On the Vibe Coders and Their Lies clerk: Auto-Summarize Your Claude Code Sessions AI Weekly — 2026/04/10–04/17 | The Model Lockdown Is Here, but the Toolchain Is the Real Battleground AI 週報 — 2026/04/10–2026/04/17 模型封鎖潮來了,但工具鏈才是真戰場 Maybe this is how Open-Source apps are born... 🚀 Fine-Tune LLMs with LoRA and QLoRA: 2026 Guide tRPC v11 + Next.js App Router: End-to-End Type Safety Without the Boilerplate ShadCN UI in 2026: Why I Stopped Installing Component Libraries and Started Owning My Components SaaS Billing in React Server Components: Stripe + Supabase Without a Single `useEffect` Join our DEV Weekend Challenge — $1,000 in Prizes Across TEN winners! Submissions Due April 20 at 6:59 AM UTC. Implementing FSRS Spaced Repetition in Flutter + Supabase — Adding Memory Science to an AI Learning App "I Texted My Localhost From the Train — Claude Code Fixed the Bug Before I Got Home" I Built a Sales Prep AI and It Went Deeper Than Expected Design to Code #2: One JSON, Eleven Outputs Solving the 100M-Row Problem: A Summary Table Pattern for High-Volume Push Notification Logs Flutter Web With Wasm: What Actually Changes For Developers I Built 50 Royalty-Free Soundtracks for My Side Project in a Weekend Using AI Music Generation The Vibe Coding Security Checklist: 7 Things to Check Before You Ship Stop Letting Googlebot Guess Fix Your React App's SEO Right Desconstruindo o Streaming do LinkedIn: Como Criar um Engine de Extração de Vídeo de Alta Performance com HLS e FFmpeg (EDA Part-1) EDA (Exploratory Data Analysis) Explained With Real Life — Why Looking at Your Data Is the Most Important Step in Machine Learning Brand Relationship Management at Scale: Our 4-Touch Outreach System for 200+ Brands Why String.fromEnvironment() Might Return an Empty String in Dart JGuardrails 1.0.0 — Hardening Java LLM Apps Against Jailbreaks, Toxicity, and Prompt Injection Plan and Schedule a Full Week of Threads Content From One Claude Conversation Coding Cat Oran Ep3, Five Tables Changed Everything Updated: BFF Pattern I'm done watching freelancers get buried by 200 proposals. So I'm building the alternative. This is my first post BFS Algorithm in Java Step by Step Tutorial with Examples Tracking LLM Pricing Monthly: An Open Dataset for 22 AI Models How We Measure Content ROI on a Comparison Site: Revenue Attribution Without Perfect Data Introducing Nova AI Ops: The AI-Native Operating System for SRE Teams I built a free desktop video downloader for Windows — Grabbit How Talkie OCR Helps Vision-Impaired & Dyslexic Users Read the World Around Them VRCFaceTracking安装和iPhone面捕配置教程,有bug Even CrowdStrike Can't See Your Agents The Automation Gold Rush: What n8n Workflows and Claude Are Opening Up for Developers Right Now
OpenCode: 160K Stars, Model-Agnostic, and It Beat Claude Code on Debugging
Anup Karanjkar · 2026-06-19 · via DEV Community

On May 6, 2026, OpenCode crossed 160,000 GitHub stars — making it the most-starred open-source AI coding agent in history, outpacing every proprietary competitor by raw community signal. It now sits at 172,000+ with 7.5 million monthly active developers. No marketing budget. No IDE lock-in. No subscription product. Just a terminal-first agent that lets you swap models mid-session without touching a config file.

That doesn't mean it beats Claude Code at everything. It doesn't. In a 38-task benchmark on a real 200KLOC TypeScript monorepo, Claude Code completed 82% of tasks vs OpenCode's 74%. On complex multi-file refactors, the gap was 8 percentage points and 9 minutes average execution time vs 16. Claude Code is faster and more accurate on architectural work.

But OpenCode's case isn't about winning every benchmark row. It's about what you give up when you don't own your AI coding stack — and which teams that trade-off actually hurts.

What OpenCode Actually Is

OpenCode is an open-source, terminal-native AI coding agent that runs as a persistent client/server pair. The server — launched with opencode serve — handles AI communication and session state in a local SQLite database. Your terminal TUI, desktop app, or IDE extension connects to it as a client. Sessions survive terminal crashes, can be accessed remotely over SSH, and support multiple simultaneous agents without duplicating model calls or splitting state.

The model-agnostic layer is the core architectural bet. OpenCode routes requests across 75+ providers: Anthropic (Claude Sonnet 4.6, Opus 4.8), OpenAI (GPT-5.5, GPT-5.6 preview), Google (Gemini 3.1), DeepSeek V4, and any local model running via Ollama. You configure the provider per session, per subagent, or per task type. A Scout subagent can hit GPT-5.5 for external research while your main coding loop runs on Claude Sonnet — without reconfiguring anything or restarting the server process.

OpenCode launched in June 2025 and hit 160K stars in under a year. The 900+ contributors who shipped that weren't optimizing for market share. They were solving a specific problem: model lock-in is a hidden cost no benchmark measures, and the tools with the most momentum in 2025 all required you to commit to one vendor's API.

The LSP Advantage Nobody Talks About

The most technically consequential feature in OpenCode isn't the model routing. It's LSP integration.

Claude Code and OpenAI Codex do not feed Language Server Protocol diagnostics into the agent loop by default. When you ask Claude Code to refactor a TypeScript function and it produces code with a type error, it doesn't know about the error unless you manually paste the compiler output or run a verification step. OpenCode auto-downloads LSP servers for each language when it detects a matching file extension, then feeds the live diagnostic stream directly to the active model during generation.

In practice this changes how the agent handles errors. Instead of generating code → running it → having you report the error → regenerating in a new turn, OpenCode receives the type error mid-generation and corrects it in the same pass. On 30+ languages including TypeScript, Python, Go, Rust, Java, and C++, the LSP loop is fully automated and requires no user configuration beyond installing the language toolchain.

One development team running OpenCode on a 400-file Go service reported a 30% reduction in edit-run-debug cycles on refactoring tasks specifically because of this pattern. That's a number that's hard to capture in a 38-task benchmark but shows up clearly in a two-week sprint retrospective.

This is also the direct explanation for OpenCode's 90% debugging task completion rate vs Claude Code's 80% in the production benchmark. Debugging is precisely the task type where live diagnostic feedback during generation makes the largest difference. An agent that knows about the compiler error while writing the fix handles it differently than one that needs a separate human-mediated feedback loop.

Three Built-In Subagents

OpenCode ships with three purpose-built subagents available on every installation. Understanding what each does changes how you structure work in practice.

General: Full tool access — reads, writes, runs commands, hits APIs. Used for the main coding loop, multi-step tasks with side effects, and anything requiring persistent state across multiple tool calls. This is what runs when you type a task with no specific prefix.

Explore: Read-only, no writes, no command execution. Designed for codebase navigation — symbol lookup, dependency tracing, call graph analysis, understanding an unfamiliar service. The constraint is the feature: read-only access means you can run it safely on production codebases, shared repos, or regulated environments where accidental writes are unacceptable. It also runs at lower cost since it doesn't need the full toolset.

Scout: Read-only access to external dependencies and documentation. When you're working with a library you just added or an SDK with thin docs, Scout can browse documentation sites, parse README files, and pull from GitHub issues without touching your codebase. Added in the 0.14 release (late May 2026), it addresses a real gap: how do you give an agent research capability without also giving it write permissions to live infrastructure? Scout answers that with a hard permission boundary.

Beyond these three, OpenCode accepts custom subagent definitions in JSON or markdown — same pattern as Claude Code's CLAUDE.md context injection system, but targeting the agent harness itself rather than just prompt context. You can define a "SecurityReviewer" subagent that runs read-only on your auth service with a specific system prompt, or a "TestWriter" that routes to a cheaper model for mechanical test generation while the main loop uses a frontier model for architecture decisions.

Background Subagents: The Async Case

Background subagents are the June 2026 feature drawing the most developer attention. They let you dispatch long-running tasks — large refactors, multi-file test generation, codebase-wide searches — without blocking your active terminal session. The background agent runs in the persistent server process, posts updates to an event log, and surfaces completion when it's done. No second terminal pane to monitor. No separate process to babysit.

The workflow looks like this: opencode bg "run tests for src/api/** and write coverage report to docs/coverage.md" queues the task, and you continue editing. The event log is accessible via opencode log at any point. Completion triggers a desktop notification.

This matters for the benchmark numbers. In the 38-task test that produced the 74% overall completion rate, OpenCode ran 23% of its tasks as background subagents. Those tasks took longer in wall-clock time, which is part of why the 16-minute average execution time was higher than Claude Code's 9 minutes. But those tasks were running in parallel with other work — the 16-minute number in isolation overstates the actual productivity cost.

The Actual Cost Model

OpenCode is MIT-licensed and free. The cost is the model API you choose to connect to it.

For teams running open-weight models on cloud GPUs: effectively zero for the tool itself. DeepSeek V4 Pro at $0.07 per million input tokens vs Claude Sonnet 4.6 at $3.00 per million is a 42x cost difference. A four-person development team running typical coding agent workloads — roughly 15–20 million input tokens per seat per month — pays ~$45/month total on DeepSeek vs ~$240/month on Sonnet API keys. Claude Code Pro at $100/seat/month runs to $400 for the same team.

OpenCode's Go tier at $10/month adds access to managed open-weight model endpoints (eliminating the need to run your own GPU), priority support, and enterprise SSO. It does not add exclusive model access — if you want Claude Sonnet at full speed, you use your own Anthropic API key regardless of tier. The Go tier is positioned at teams who want the cost efficiency of open-weight models without the infrastructure overhead of self-hosting.

For solo developers at typical usage levels, Claude Code's flat $100/month subscription frequently undercuts per-token API costs when you're hitting the model hard. The cost case for OpenCode is strongest for teams, for users who want open-weight model quality (which has narrowed substantially vs frontier models in 2026), and for air-gapped deployments where API calls to Anthropic or OpenAI are architecturally excluded.

Benchmark Reality Check

The 38-task production test used a real TypeScript monorepo, not SWE-bench or Terminal-Bench eval sets. Tasks: 12 complex refactors, 10 debugging sessions, 9 test generation runs, 7 documentation tasks.

Task type Claude Code OpenCode Winner
| Complex refactors | 83% | 67% | Claude Code (+16pp) |

| Debugging sessions | 80% | 90% | OpenCode (+10pp) |

| Test generation | 78% (73 tests written) | 78% (94 tests written) | Tie (OpenCode more thorough) |

| Documentation | 71% | 86% | OpenCode (+15pp) |

| Overall | 82% | 74% | Claude Code (+8pp) |

The debugging edge is explained directly by LSP. The documentation edge is less obvious — both tools wrote from the same codebase. The difference appears to be OpenCode's thoroughness optimization: it ran the full existing test suite (200+ tests) before writing documentation claims about behavior, while Claude Code verified only the specific functions being documented. Both approaches are valid; OpenCode's just produces fewer documentation inaccuracies on codebases where behavior diverges from expectations.

A separate AlterSquare 50-task production test found Claude Code introduced more technical debt in its solutions — specifically more subset testing (verifying only the changed code, not the full regression surface) and more architectural shortcuts under time pressure. OpenCode's slower average completion time was correlated with fewer follow-up fix tasks in the two weeks after the initial run. That doesn't show up as "OpenCode won" in completion rate. It shows up in sprint velocity two weeks later.

The Full Stack Decision Table

Tool Price Model lock-in LSP in loop Background agents Strongest at
| OpenCode | Free / $10/mo Go | None (75+ providers) | Yes, auto-configured | Yes (v0.14+) | Debugging, docs, air-gapped, cost-sensitive teams |

| Claude Code | $100/mo Pro | Anthropic only | Partial (not default loop) | No | Complex refactors, speed, GitHub ecosystem |

| Codex CLI (OpenAI) | Pay-per-token | OpenAI only | No | No | Terminal-Bench 2.1 score (83.4%), OpenAI integrations |

| Cursor | $20/mo Business | Multi-model, IDE-locked | Via IDE | Beta | IDE users, inline autocomplete speed, enterprise SSO |




When Not to Switch

The 7-minute average execution time gap (9 vs 16 minutes on identical refactoring tasks) adds up on teams measuring sprint velocity. If your definition of done is "passed CI and merged in the same session," Claude Code's speed advantage is real and consistent.

Claude Code also owns 10%+ of all public GitHub commits, peaked at 326,000 commits per day in March 2026, and has deep integrations with GitHub Actions, Copilot, and Anthropic's managed agents platform. If you've built custom workflows on Claude Code's skills architecture — hooks, MCP integrations, CLAUDE.md-driven subagents — the switching cost is non-trivial. OpenCode's custom subagent definitions in JSON/markdown are functionally equivalent for many use cases, but migrating an existing toolkit takes real time.

For teams that have never been on Claude Code and are evaluating from scratch in June 2026: OpenCode's model-agnostic design means you can start with the Anthropic API and switch to DeepSeek V4 when cost pressure hits. You don't have to commit the architecture to a single vendor at setup time. That optionality has a real value that doesn't appear in any benchmark table.

Three Concrete Steps

Run OpenCode alongside your current tool for two weeks on debugging and documentation tasks specifically. Those are the task types where the LSP loop and thoroughness advantage are most measurable, and they're tasks most engineering teams do daily without treating them as evaluation surfaces.

If you have four or more developers at $100/seat/month on Claude Code, run the API cost math with a DeepSeek V4 endpoint. At $0.07 per million input tokens vs $3.00, the breakeven on a managed GPU instance is around 2,500 coding sessions per month. Most active four-person teams clear that threshold.

OpenCode's 172K stars (as of June 19, 2026) vs Claude Code's 326K daily GitHub commits is evidence these tools aren't mutually exclusive. A growing pattern in production teams is running Claude Code for complex architectural work and OpenCode for debugging, test generation, and documentation — same codebase, model routing by task type rather than tool loyalty. That's exactly what OpenCode's multi-provider architecture was built for.

Originally published at wowhow.cloud