惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MyScale Blog
MyScale Blog
Apple Machine Learning Research
Apple Machine Learning Research
H
Help Net Security
雷峰网
雷峰网
V
Visual Studio Blog
G
Google Developers Blog
Microsoft Azure Blog
Microsoft Azure Blog
Hugging Face - Blog
Hugging Face - Blog
爱范儿
爱范儿
IT之家
IT之家
Engineering at Meta
Engineering at Meta
Microsoft Security Blog
Microsoft Security Blog
aimingoo的专栏
aimingoo的专栏
大猫的无限游戏
大猫的无限游戏
M
MIT News - Artificial intelligence
月光博客
月光博客
A
About on SuperTechFans
B
Blog RSS Feed
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
The GitHub Blog
The GitHub Blog
N
Netflix TechBlog - Medium
J
Java Code Geeks
云风的 BLOG
云风的 BLOG
Blog — PlanetScale
Blog — PlanetScale

Codeium Blog Posts

Opus 4.7 (fast mode) is now available in Windsurf | Devin Fast and Comprehensive Code Review, Now in Windsurf Windsurf 2.0: Introducing the Agent Command Center and Devin in Windsurf Introducing our new Windsurf pricing plans GPT-5.4 is now available in Windsurf Gemini 3.1 Pro is now available in Windsurf | Devin Claude Sonnet 4.6 is now available in Windsurf | Devin GLM-5 and Minimax M2.5 are now available in Windsurf | Devin GPT-5.3-Codex-Spark is now available in Windsurf | Devin Windsurf Arena Mode Leaderboard: The People Want Speed Opus 4.6 (fast mode) is now available in Windsurf | Devin Opus 4.6 is now available in Windsurf | Devin Windsurf Tab v2: 25-75% more accepted code with Variable Aggression Wave 14: Arena Mode - May the Best Model Win GPT-5.2-Codex is now available in Windsurf! Windsurf Wave 13: Merry Shipmas GPT 5.2 is now available in Windsurf! Opus 4.5 is now available in Windsurf | Devin GPT 5.1, GPT 5.1-Codex, and GPT-5.1-Codex Mini are now available in Windsurf Introducing SWE-1.5: Our Fast Agent Model Cognition and Windsurf Cognition (Windsurf) Named a Leader in the 2025 Gartner® Magic Quadrant™ for AI Code Assistants Windsurf Queued Messages Release Windsurf Wave 12: Devin features in Windsurf Wave 11: Just Keep Shipping Our Commitment to Windsurf The Next Chapter The Next Stage of Windsurf Changelist: June 2025 Windsurf and AHEAD Form Strategic AI DevOps Partnership
Introducing Adaptive: a smarter way to use Windsurf
2026-04-06 · via Codeium Blog Posts

We’re launching multiple updates to Windsurf today: an Adaptive model router, a redesigned model picker with pricing context, and the removal of daily limits for Max.

We’ve heard clear feedback that our new pricing plans were too opaque and, in some cases, too restrictive. With this update, we’re aiming to address that directly by giving users better visibility into model costs and making it easier to manage your quota.

Introducing our Adaptive model router

We’re rolling out an adaptive model router. Adaptive is designed to intelligently select the best models for your tasks. By automatically choosing the right model for each task and avoiding overuse of premium models, Adaptive will help you make your quota last longer.

You can use our new model router by choosing Adaptive in the model picker.

Adaptive selected in the Windsurf model picker

When you choose the adaptive model, we will dynamically choose the right underlying model for your task—while drawing down your quota at a fixed per-token rate. To start, we are including generous resource limits with your quota and offering extra usage beyond your quota at USD 0.50 per 1M input tokens, USD 2.00 per 1M output tokens, and USD 0.10 per 1M cache read tokens for the next 2 weeks.

We’re rolling out the Adaptive model to all self-serve users today, including Pro, Max, and Teams. We expect this to be the best default option for most users who want to stay within the quota limits throughout the month, and we’ll continue deploying new updates to improve routing performance.

Updated model picker with pricing context

When we announced new pricing, we should have been clearer that your quota and extra usage will be billed based on how many tokens your requests consume. To address that, we’re introducing a new model picker design that shows token pricing information directly, which is the exact rate extra usage is billed at.

Updated model picker showing token pricing for Claude Opus 4.6

As you’ll notice from the new model picker design, prompt caching is a critical component to request costs. To make this more obvious, we’ve integrated a prompt cache timer directly into the context window indicator so you can track it more closely.

Prompt cache timer integrated into the context window indicator

Finally, we have updated the response cards after messages to include token counts so you can understand exactly how the message cost was calculated.

No more daily limits for Max users

While the average user hasn’t been heavily impacted by the new quotas, we should have been clearer that for the heaviest users—who learned to squeeze the most out of each prompt—the new token-based pricing is a significant change.

The new Max plan is designed for power users who want to drive AI to its fullest, but we’ve heard your feedback that the daily limits felt too restrictive for bursty work.

Therefore, starting today, Max users will no longer have a daily quota. Your weekly limit remains, but you’re free to use it however your workflow demands.

We also considered removing the daily limits for all plans, but after analyzing the data decided it’s too easy to exhaust an entire week’s quota before you get a hang of prompt caching and model selection. The daily quota gives you a safety net to continue for free the next day—if you want to continue coding immediately once your quota is exhausted, you can always purchase extra usage or upgrade to Max.

What we’re working on next

Between Adaptive, the updated model picker, and the removal of daily limits for Max, we think this is a meaningful step toward a more flexible and transparent Windsurf. All updates are live today: download the latest version of Windsurf to try them out.

But we’re not stopping there: we want to continue to improve performance and efficiency for all Windsurf users. That’s why we’re working on a new, more efficient harness that will intelligently incorporate a multi-model architecture and subagents to deliver higher quality outputs at a lower overall cost. We plan to share more soon about this.