惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

I
InfoQ
D
DataBreaches.Net
Engineering at Meta
Engineering at Meta
GbyAI
GbyAI
Martin Fowler
Martin Fowler
Security Latest
Security Latest
Cisco Talos Blog
Cisco Talos Blog
MongoDB | Blog
MongoDB | Blog
D
Darknet – Hacking Tools, Hacker News & Cyber Security
IT之家
IT之家
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
S
Security Affairs
www.infosecurity-magazine.com
www.infosecurity-magazine.com
博客园_首页
L
LINUX DO - 最新话题
Know Your Adversary
Know Your Adversary
S
Schneier on Security
The Last Watchdog
The Last Watchdog
Attack and Defense Labs
Attack and Defense Labs
T
Tenable Blog
G
GRAHAM CLULEY
Y
Y Combinator Blog
P
Palo Alto Networks Blog
L
LINUX DO - 热门话题
Hugging Face - Blog
Hugging Face - Blog
W
WeLiveSecurity
C
Cybersecurity and Infrastructure Security Agency CISA
aimingoo的专栏
aimingoo的专栏
博客园 - 司徒正美
The Register - Security
The Register - Security
T
The Exploit Database - CXSecurity.com
MyScale Blog
MyScale Blog
M
MIT News - Artificial intelligence
Cyberwarzone
Cyberwarzone
雷峰网
雷峰网
T
Tailwind CSS Blog
V2EX - 技术
V2EX - 技术
T
Threat Research - Cisco Blogs
S
Secure Thoughts
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
O
OpenAI News
C
Cyber Attacks, Cyber Crime and Cyber Security
The Cloudflare Blog
量子位
Apple Machine Learning Research
Apple Machine Learning Research
T
Threatpost
S
SegmentFault 最新的问题
小众软件
小众软件
Google DeepMind News
Google DeepMind News
Help Net Security
Help Net Security

Codeium Blog Posts

Opus 4.7 (fast mode) is now available in Windsurf Fast and Comprehensive Code Review, Now in Windsurf Windsurf 2.0: Introducing the Agent Command Center and Devin in Windsurf Introducing Adaptive: a smarter way to use Windsurf Introducing our new Windsurf pricing plans GPT-5.4 is now available in Windsurf Gemini 3.1 Pro is now available in Windsurf Claude Sonnet 4.6 is now available in Windsurf GLM-5 and Minimax M2.5 are now available in Windsurf GPT-5.3-Codex-Spark is now available in Windsurf Windsurf Arena Mode Leaderboard: The People Want Speed Opus 4.6 (fast mode) is now available in Windsurf Opus 4.6 is now available in Windsurf Windsurf Tab v2: 25-75% more accepted code with Variable Aggression Wave 14: Arena Mode - May the Best Model Win GPT-5.2-Codex is now available in Windsurf! Windsurf Wave 13: Merry Shipmas GPT 5.2 is now available in Windsurf! Opus 4.5 is now available in Windsurf GPT 5.1, GPT 5.1-Codex, and GPT-5.1-Codex Mini are now available in Windsurf Introducing SWE-1.5: Our Fast Agent Model Cognition and Windsurf Cognition (Windsurf) Named a Leader in the 2025 Gartner® Magic Quadrant™ for AI Code Assistants Windsurf Queued Messages Release Windsurf Wave 12: Devin features in Windsurf Wave 11: Just Keep Shipping Our Commitment to Windsurf The Next Chapter The Next Stage of Windsurf Changelist: June 2025 Windsurf and AHEAD Form Strategic AI DevOps Partnership Our Brand Percentage of Code Written Wave 10: Other Announcements Wave 10: The Windsurf Browser Wave 10: Planning Mode athenahealth Advances Healthcare Innovation with Windsurf Changelist: May 2025 Statement on Anthropic Model Availability An Inflection Point for U.S. Government Windsurf Fuels Mercado Libre’s Growth Strategy SWE-1: Our First Frontier Models Wave 8: UX Features + Plugins Update Wave 8: Cascade Customization Features Windsurf Wave 8: Teams & Enterprise Features Changelist: April 2025 An Update to Our Free Plan An Update to Our Pricing Changelist: March 2025 Windsurf Named 2025’s Forbes AI 50 Recipient Windsurf Wave 7 Forge Deprecation
Self-Hosted Deployment Maintenance Mode
2025-05-12 · via Codeium Blog Posts

Cognition5 min read

tl;dr We have decided to place our self-hosted offering in maintenance mode, and offer a new single-tenant hosted offering that can expose our agentic capabilities.

A little bit of history and context first.

Two years ago, we released the first enterprise offering we ever had, which was a self-hosted deployment. We released this even before we had a hosted cloud solution. Why? Well, our background as a GPU infrastructure company meant that we could uniquely build such an offering, and at a time when that was the entirety of our expertise, we saw an opportunity to work with large enterprises that were in regulated industries or had other security-related constraints. It was also early in the generative AI era, and it would have been harder at the time, as a tiny startup with essentially zero security certifications, to gain much traction with a hosted solution, even with enterprises that weren’t strictly regulated. It made sense, and we were, by most metrics, highly successful with this offering.

However, two years in the generative AI space might as well be two decades in normal software development. A number of different factors have progressed simultaneously:

  • We became a lot better at creating offerings that were intermediate between fully commercial-cloud hosted and fully self-hosted, and our security posture massively improved (not to mention size and stability of the company). We created a Hybrid deployment to solve the persisted data concerns, we became the first AI code assistance platform to receive FedRAMP High authorization to support highly regulated industries (not just SOC2 Type 2!), and we launched a cluster in Germany to satisfy European regulations around data residency.
  • The underlying technology continued to progress. The frontier models of today dwarf the largest foundation models two years ago in size, context length, and practically any other metric you can think of. It has come to the point that hosting even our best Tab model would require a full rack of frontier GPUs to serve, and those are our smaller models. This stops making sense from a TCO perspective for almost all companies. Connecting to a third-party endpoint for models from the foundation labs has essentially become a requirement for our self-hosted deployment. And as users try these tools in their spare time on personal machines, they start noticing the delta in performance from using smaller, previous generation models under the hood.
  • The fork in capabilities that we could offer on self-hosted and hybrid/cloud grew. For quite a while, even if it was with different sized models, we could provide similar sets of capabilities between our hosted Cloud/Hybrid solution and our self-hosted offering; However, with the advent of Cascade and the new cutting edge agentic capabilities, which simply cannot be run entirely in a self-hosted environment, there is a very meaningful delta in what we can offer our customers between our different deployments. We no longer can say honestly that we are providing our self-hosted customers even close to the best of what we have to offer.
  • Companies have become more comfortable with Cloud/Hybrid solutions that provide guarantees around zero-data retention. Maybe it is just time and familiarity with the technology, maybe it is our new certifications and more stability as a company, or maybe it is an understanding of the kinds of agentic capabilities that are restricted to these offerings. Likely some combination of all three.

Combining all of these together, we had to make a tough decision: in a world where self-hosted will, within time, not be the answer for over 99% of developers, do we (a) continue to invest in building on the self-hosted deployment knowing that we will likely lose almost all such customers over time, or (b) double down all of our resources to maximize the value we can drive to the vast majority of developers at the vast majority of our current and future customers.

While we stuck with (a) for a while, partially because a meaningful chunk of our customers have trusted us with that offering over the last couple of years, we eventually came to the realization that we would actually be doing our enterprise customers a disservice if we did not adapt to the times and adopt option (b). Given our certifications, security posture, GTM and partner motions, and product-level bets (see Wave 8), we hope that it is still clear that we are still primarily focused on the enterprise.

And we are not stopping in iterating on our security posture. Today, we are also simultaneously rolling out a single-tenant hosted offering as an additional security guarantee to our largest customers. This does require a minimum commitment given the hardware costs on our end, but we recognize that multitenancy is the second most common objection to Cloud solutions after data retention.

By maintenance mode, we will of course continue to support our current self-hosted customers until the end of their term and provide great financial incentives to switch to and adopt our other deployment offerings, but we are no longer investing in feature development or bringing on new customers to the self-hosted platform.

We will also never say never. Maybe one day in the future, there are new changes in the tech in a way where self-hosting becomes possible and more meaningfully equivalent to the hosted deployment options. Our customers are not just betting on our offerings today, but on the ability of our team to adapt to the rapidly-changing landscape of this space to consistently deliver the frontier.

We are grateful for the success that the self-hosted offering enabled us up to today, but are even more excited for what the future holds.