惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

A
Arctic Wolf
T
Tenable Blog
T
Troy Hunt's Blog
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
P
Privacy & Cybersecurity Law Blog
NISL@THU
NISL@THU
Application and Cybersecurity Blog
Application and Cybersecurity Blog
H
Hacker News: Front Page
S
Secure Thoughts
AWS News Blog
AWS News Blog
L
LINUX DO - 最新话题
D
Darknet – Hacking Tools, Hacker News & Cyber Security
M
MIT News - Artificial intelligence
T
Tor Project blog
S
Schneier on Security
PCI Perspectives
PCI Perspectives
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
美团技术团队
Google DeepMind News
Google DeepMind News
V
Visual Studio Blog
爱范儿
爱范儿
Google DeepMind News
Google DeepMind News
Cyberwarzone
Cyberwarzone
T
The Exploit Database - CXSecurity.com
罗磊的独立博客
T
Threat Research - Cisco Blogs
Recent Commits to openclaw:main
Recent Commits to openclaw:main
V
V2EX
C
CXSECURITY Database RSS Feed - CXSecurity.com
Stack Overflow Blog
Stack Overflow Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
G
GRAHAM CLULEY
L
LINUX DO - 热门话题
D
Docker
J
Java Code Geeks
GbyAI
GbyAI
H
Heimdal Security Blog
The Hacker News
The Hacker News
MongoDB | Blog
MongoDB | Blog
V
Vulnerabilities – Threatpost
T
Tailwind CSS Blog
Cloudbric
Cloudbric
TaoSecurity Blog
TaoSecurity Blog
C
CERT Recently Published Vulnerability Notes
Y
Y Combinator Blog
Recorded Future
Recorded Future
Cisco Talos Blog
Cisco Talos Blog
T
Threatpost
The Register - Security
The Register - Security
Hacker News - Newest:
Hacker News - Newest: "LLM"

OfficeChai

These Are The 10 Cheapest AI Models In The World [June 2026] 18 Best AI Tools For English Speaking (With Examples) [2026] AI Impact? Vacancy Rates For US Office Properties Are Now Highest Since The 2008 Crisis KPMG Pulls Report Praising AI After It Was Found To Have Fake AI-Generated Citations India's Sarvam Raises $234 Million At $1.5 Billion Valuation After SpaceX Stock Pops 20%, Musk Has Made More Money In The Last 24 Hours Than Warren Buffett Made In His Entire Career OfficeChai Nobody Is Using AI Better Than Meta: NVIDIA CEO Jensen Huang 21 Best AI Tools For Animation (With Examples) [2026] 22 Best AI Tools For Architecture (With Examples) [2026] Datacenter Construction Spending Has Eclipsed Public Transportation Spending In The US China Scraps 12,000 Degree Courses, Mainly In Arts And Humanities, To Prepare For AI Age OfficeChai There Is No Job Loss With AI: David Friedberg Loop Between Human Capital And "Token Capital" Will Be The New IP For Firms, Says Satya Nadella How to Reduce Dependency on Key Employees 8 Google Index Checker Use Cases Beyond New Blog Posts Memory Squeeze? Smartphone Purchases Are Down Globally 21 Best AI Tools For Accounting (With Examples) [2026] AI For Voice Generation: 22 Best Options (With Examples) [2026] These Are The Most Popular Image Generation Models On OpenRouter [June 2026] Search Traffic For Websites Is Down 25% Over The Last Year Because Of AI: a16z Data Agentic Coding Has Led To A 50% Increase In Number Of Apps, But Most Are Finding Very Few Users: SimilarWeb Data OpenRouter Launches Fusion API, Which Uses A Combination Of Models To Achieve Fable-Like Performance At Half The Price Dario Amodei Refused To De-Deploy Or Fix Vulnerabilities In Fable Before US Export Controls, Says David Sacks 23 Best AI Tools For Notes Making (With Examples) [2026] 16 Best AI Tools For Astrology (With Examples) [2026] How Jensen Huang Once Had To Ask SEGA's CEO To Pay NVIDIA For A Technology That Didn't Work ChatGPT Already Has 11% Of The Search Market: OpenAI CFO Sarah Friar SpaceX Has Now Launched More Satellites Than Rest Of Humanity Combined Across History Globalization Is Dead, Time For India To Wake Up Says Sridhar Vembu After US Bans Anthropic Mythos And Fable Models For Foreign Users Elon Musk Becomes World's First Trillionaire After Record SpaceX IPO Anthropic Suspends Access To Mythos And Fable Models Following US Govt Directive Against Foreign Users 27 Best AI Tools For Market Research (With Examples) [2026] Why Jeff Bezos Makes Important Decisions Early in The Morning Education And Healthcare IT Have Been The Hardest Areas To Invest In: Peter Thiel Giving AI Long-Term Goals Could Lead To The Emergence Of Self-Preservation: Geoffrey Hinton Your Startup Doesn't Have a Hardware Problem. It Has an Accountability Problem Cyber Incidents Rarely Start With a Hacker: The Weak Links Businesses Overlook What Makes an App Worth Returning to Every Day? 21 Best AI Tools For Lead Generation (With Examples) [2026] How NBA Player Shaquille O'Neal Became An Early Investor In Ring AI For Kids Learning: 22 Best Options (With Examples) [2026] These Are The Most Popular AI Model Companies On OpenRouter [June 2026] Advanced Fintech and NeoBank Software Development Solutions: Building the Digital Banks of Tomorrow TRON Payments: Integrating AML Checks Into Business Workflows 18 Best AI Tools For Resume (With Examples) [2026] 16 Best AI Tools For UI Design (With Examples) [2026] These Are Top 10 Countries Generating The Most Internet Traffic How to Choose the Best Magento Agency for Your Store These Are The Best AI Models For Creative Writing [June 2026] AI For Managers: 28 Best Tools (With Examples) [2026] 17 AI Tools For Trading (With Examples) [2026] AI Has Led To An Explosion Of New Apps, But Nearly None Have Managed To Garner Significant Usage Cloudflare CEO Matthew Prince Says Vinod Khosla Asked Him To Fire His Co-founders For Him To Invest In His Company Australia’s AirTrunk To Invest $30 Billion To Develop Datacenters In India Anthropic Says That Their Employees Are Using AI To Write 8x More Code Compared To 18 Months Ago Anthropic Is Extremely Expensive, Many Are Urgently Looking For Alternatives: Microsoft AI CEO Mustafa Suleyman Sergey Tokarev on creating DIY “Beehives” and a free guidebook AI Crypto Price Prediction: How Accurate Are Machine Learning Models? Why Anthropic Could Find It Hard To Maintain Its $965 Billion Valuation Startup CEO Says They're Saving "Millions Of Dollars" By Replacing Anthropic Models With DeepSeek Ola Cabs' Valuation Falls 99% From Peak, Now Valued At Just $70 Million By Vanguard After TCS Case, Former Wipro Employee Alleges Attempt At Religious Conversion By Coworkers Bot Traffic Has Surpassed Human Traffic On The Internet For The First Time In History, Clouflare Says ChatGPT's Free Users Do 7 Queries Per Day, Those On $20 Plan Do 3x More: CFO Sarah Friar How Keith Rabois Had Been "Highly Skeptical" In 2023 That Anthropic Would Be Worth More Than $5 Billion In 10 Years How to Install AdGuard Home with Docker Step by Step We're Running Out Of Training Data, But Not Too Worried Because There Are Alternate Approaches: Google's Jeff Dean JioHotstar Is Hiring For 75 AI Roles Amid AI Content Push NVIDIA's Nemotron 3 Becomes Most Intelligent Open Weights Model From The US Hackers Allegedly Fooled Meta's AI To Take Over Accounts By Simply Asking It To Change User Emails Manchester Super Giants' AI Promotional Video Gets Panned As "Slop" For Glaring Cricketing Errors AI Reducing Jobs Is "Complete Nonsense": NVIDIA CEO Jensen Huang MiniMax Releases MiniMax M3, Is Competitive With Frontier Models On Many Benchmarks IIT Delhi-Incubated BotLab Dynamics Lights Up Skies With Lord Shiva Themed Drone Show During IPL Final NVIDIA Introduces RTX Spark, A New Chip Optimized For AI Agents For Windows Laptops And PCs NVIDIA Introduces Vera, A New CPU Chip For AI Agents That Is 80% Faster Than x86 CPUs OpenAI's Codex Reaches 5 Million Users, Resets Rate Limits For Users Key Factors That Influence Personal Loan Approval in India AI Is Allowing Me To Experiment And Try Crazier Things: Mathematician Terrance Tao Efficiency Of Human Learning Is Still A Thousand Times Better Than LLM Learning, Need Algorithmic Advances To Improve It: Jeff Dean San Francisco Home's Zillow Listing Says It'll Accept OpenAI Or Anthropic Stock As Payment Open-Source Models Currently Lag Proprietary Models By Just 4 Months: Epoch AI Self-Improvement Possible In AI Models Within A Year, Say Google's Top AI Leaders Digital Minds: Preparing for a Moral Challenge Before It Arrives Nearly 30% Of US-Based Y-Combinator Founders Are Of Indian Origin: SF Chronicle Data "A New Era Of PC": NVIDIA, Microsoft Windows Tease New Collaboration At Least 146,000 AI Hallucinated Citations In Papers Published In 2025, Finds Paper AI Doesn't Undergo Experiences, Has No Moral Conscience: Pope Leo XIV Claude Opus 4.8 Tops Artificial Analysis Intelligence Index, Edges Out GPT 5.5 With Score Of 61.4 Anthropic Says Its Annual Revenue Run-rate Has Now Touched $47 Billion Anthropic Raises $65 Billion At $965 Billion Valuation, Is Now Worth More Than OpenAI Claude Opus 4.8 Is Better Than Opus 4.7 But Not As Good As Mythos Preview, Says Anthropic Claude Opus 4.8 Beats GPT 5.5 On GDPval-AA Benchmark For Real World Tasks Anthropic Releases Claude Opus 4.8, Beats Opus 4.7, GPT-5.5 On Many Benchmarks GTM for Tech Startups Explained How to Use an AI Picture Generator to Create Professional Images Anthropic Is Now Generating 35% More Revenue Than OpenAI: The Information SK Hynix, Micron Join $1 Trillion Club Following AI-Led Memory Shortages
Sakana AI Launches Sakana Fugu That Matches Fable And Mythos On Some Benchmarks By Coordinating And Orchestrating Multiple Models
OfficeChai Team · 2026-06-22 · via OfficeChai

While frontier labs are competing on building the best AI models, smaller startups are looking to match their performance through innovative approaches.

Sakana AI — the Tokyo-based lab valued at $2.65 billion after its November 2025 Series B — has launched Sakana Fugu, a system that aims to match frontier-level AI performance through a fundamentally different approach: rather than training a single powerful model, it coordinates and orchestrates a pool of existing models to tackle complex tasks.

The tagline is “One Model to Command Them All,” and the core idea is that a well-orchestrated team of models can outperform any individual one. Instead of assigning fixed roles or workflows to specific models upfront, Fugu learns to dynamically assemble agents from a pool and route work between them in patterns that aren’t obvious but are, according to Sakana, highly efficient. The result is delivered through a single OpenAI-compatible API, so users aren’t managing multiple model integrations.

Sakana’s benchmark results are hard to dismiss. Across a suite of coding, reasoning, scientific, and agentic evaluations, Fugu and Fugu Ultra land near the top of the field. On SWE-Bench Pro — a demanding software engineering benchmark — Fugu Ultra scores 73.7, ahead of Claude Opus 4.8’s 69.2 and GPT-5.5’s 58.6. On LiveCodeBench, Fugu scores 92.9 and Fugu Ultra 93.2, both ahead of Gemini 3.1 Pro’s 88.5. On Humanity’s Last Exam, one of the hardest general-knowledge benchmarks available, Fugu Ultra reaches 50.0, essentially matching Opus 4.8’s 49.8.

Sakana is careful to note that neither Claude Fable 5 nor Claude Mythos Preview — Anthropic’s highest-tier models — are in Fugu’s agent pool, as neither is publicly accessible. The company says Fugu sits shoulder-to-shoulder with those models on several benchmarks, which would place it at or near the very top of what’s currently available.

The qualitative results are interesting too. In an AutoResearch experiment where an AI agent autonomously ran 123 training experiments over 14 hours on a single H100 GPU to improve a small language model’s training recipe, Fugu Ultra achieved the best mean validation score across all seeds, ahead of three frontier model baselines. An industry researcher using the system for patent landscape analysis across roughly 20 papers and several patents reported completing in a few hours a task that would normally take three to four days.

The Architecture Behind It

Fugu is grounded in two papers accepted at ICLR 2026. The first, TRINITY, uses a lightweight evolved coordinator that assigns models to Thinker, Worker, or Verifier roles across multiple turns, adapting dynamically to the task. The second, the Conductor, uses reinforcement learning to discover natural-language coordination strategies — essentially training the system to figure out how to prompt and route agents for maximum performance, rather than having engineers design those workflows by hand.

This research-first approach is consistent with Sakana’s broader identity. The company, co-founded by David Ha (formerly of Google Brain and Stability AI) and Llion Jones (co-author of the seminal “Attention Is All You Need” paper), has consistently pursued alternatives to brute-force compute scaling. Earlier this year, its AI Scientist system became the first AI to have a fully generated paper pass peer review at a machine learning conference — a milestone that landed in Nature in March 2026. Fugu represents a similar philosophy applied to model deployment: extract more from what already exists rather than building from scratch.

Sakana Fugu Pricing: Two Tiers, One API

Fugu comes in two versions. The standard Fugu model is positioned for everyday work — coding, code review, responsive chatbot services — and balances performance with low latency. Fugu Ultra is the heavier-duty option, designed for long-horizon tasks like Kaggle competitions, paper reproduction, cybersecurity assessments, and patent and literature investigations. Early users have described running Fugu Ultra through full security assessments end-to-end, including reconnaissance, vulnerability checks, and report generation, from a single instruction.

Pricing for Fugu follows an interesting structure: when only one agent is active, users pay the standard rate for that underlying model. When multiple agents coordinate, Sakana charges a single rate based on the top-tier model involved rather than stacking fees. Fugu Ultra has fixed pricing at $5 per million input tokens and $30 per million output tokens, doubling for contexts above 272K tokens.

A subscription plan is also available, with tiers at $20, $100, and $200 per month depending on usage volume. Anyone who subscribes before the end of July 2026 gets a free second month at their initial tier. The API is currently unavailable in the EU and EEA while Sakana works toward GDPR compliance.

Vendor Flexibility as a Feature

One aspect of Fugu that enterprise buyers will likely find attractive is the ability to control which models participate in the pool. Teams with data residency requirements, compliance constraints, or specific vendor preferences can exclude certain providers or models entirely. As the frontier model competition among Anthropic, OpenAI, and Google continues to accelerate — with the top models separated by only a few benchmark points — the ability to access frontier-level performance without locking into a single vendor starts to look genuinely compelling.

Sakana’s pitch is that collective intelligence, properly orchestrated, can outperform any single model — and that a startup in Tokyo, working with different constraints and different assumptions than Silicon Valley’s compute-heavy labs, might be well-placed to build it.