惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Hugging Face - Blog
Hugging Face - Blog
腾讯CDC
阮一峰的网络日志
阮一峰的网络日志
博客园_首页
Last Week in AI
Last Week in AI
月光博客
月光博客
D
DataBreaches.Net
WordPress大学
WordPress大学
雷峰网
雷峰网
酷 壳 – CoolShell
酷 壳 – CoolShell
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
博客园 - 叶小钗
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
U
Unit 42
Recent Announcements
Recent Announcements
宝玉的分享
宝玉的分享
MyScale Blog
MyScale Blog
C
Check Point Blog
F
Fortinet All Blogs
B
Blog
小众软件
小众软件
Vercel News
Vercel News
罗磊的独立博客
有赞技术团队
有赞技术团队

TestingCatalog

SpaceXAI gearing up for upcoming Grok 4.5 release ByteDance set to launch Seedance 2.5 with 3-minute output Meta prepares Scheduled tasks for Meta AI users on web Google tests new Gemini Inbox section for Workspace triage OpenAI might be preparing GPT-5.6 for next week's release Mistral releases Leanstral 1.5 model for proof engineering xAI debuts Grok Voice Agent Builder for Enterprises Condense launches proxy to cut AI coding agent bills by 66% Vellum adds agent-to-agent AI collaboration for Slack Early look at Anthropic's Claude Science app for researchers Google launches Nano Banana 2 Lite and Gemini Omni Flash Google might be testing Gemini Flash upgrade on LM Arena Anthropic may impose KYC restrictions for Fable 5 access Anthropic launches Claude Sonnet 5 model on Claude and APIs NoimosAI launches Creative Agent for brand assets Apify lets AI Agents pay via Coinbase x402 for web tools Bloome launches chat platform for AI agent teams Meituan launches LongCat-2.0 1.6T parameter model Cursor releases its iOS app for vibe coding on the go OpenAI prepares upgraded Office controls for Codex OpenAI tests gifting Codex credits as new growth strategy Microsoft launches MAI-Code-1-Flash on GitHub Copilot Google adds Computer Use to Gemini 3.5 Flash Google tests notebook collections for NotebookLM Microsoft adds Copilot finance tools to Excel for M365 users DeepReinforce releases Ornith-1.0 open-source coding models Gemini to get voice dictation and Magic Pointer on desktop Meta launches AI glasses with three new styles from $299 Anthropic launches Claude Tag on Team and Enterprise plans Mistral launches OCR 4 for multilingual document extraction
OpenAI launches GPT-5.6 Sol preview for select partners
https://www.facebook.com/testingcatalog · 2026-06-27 · via TestingCatalog

OpenAI is opening a limited preview of GPT-5.6, led by Sol, its new flagship model, alongside Terra for lower-cost everyday work and Luna for faster, cheaper workloads. The preview starts with a small group of trusted partners, with access initially through the API and Codex, while broader access for ChatGPT, Codex, and API users is planned in the coming weeks.

Good new first: Sol is a smart, efficient, and a significant step forward. It is the same price as GPT-5.5. Also launching in the GPT-5.6 family is Terra, with 5.5-level performance at half the price.

Bad news: at the request of the US government, it is launching today in…

— Sam Altman (@sama) June 26, 2026

GPT-5.6 Sol arrives as OpenAI’s strongest model in the new family, with gains positioned around agentic coding, biology workflows, and cybersecurity tasks. The model adds a new max reasoning effort for deeper problem solving and an ultra mode that uses subagents to work on complex tasks beyond a single-agent setup. OpenAI says Sol sets a new state of the art on Terminal-Bench 2.1 and shows stronger GeneBench v1 results than GPT-5.5 while using fewer tokens.

OpenAI

The release is being handled as a controlled rollout because of the model’s cyber and biological capabilities. OpenAI says Sol, Terra, and Luna are classified as High capability in both Cybersecurity and Biological and Chemical risk under its Preparedness Framework, while not reaching the High threshold for AI self-improvement or the Cyber Critical threshold. In browser exploit tests involving Chromium and Firefox, Sol identified bugs and exploitation primitives but did not autonomously produce a full-chain exploit under the tested conditions.

OpenAI is pairing the model upgrade with a layered safeguard stack covering:

  1. Model-level refusal behavior
  2. Real-time cyber and biology misuse classifiers
  3. Account-level review
  4. Differentiated access
  5. Monitoring, enforcement, and ongoing testing

Some preview users may see blocked requests or slower responses when generation is paused for extra review, especially in dual-use security contexts where defensive and offensive work can initially look similar.

The company says it used more than 700,000 A100-equivalent GPU hours for automated red teaming focused on universal jailbreaks, alongside human expert and third-party testing. OpenAI plans to keep testing during the preview period and publish an updated system card when the GPT-5.6 family moves toward general availability.

Pricing starts at $5 per 1 million input tokens and $30 per 1 million output tokens for Sol, $2.50 input and $15 output for Terra, and $1 input and $6 output for Luna. GPT-5.6 also adds explicit cache breakpoints, a 30-minute minimum cache life, cache writes at 1.25x the uncached input rate, and cache reads with a 90% cached-input discount. OpenAI also plans to launch GPT-5.6 Sol on Cerebras in July at up to 750 tokens per second for select customers.

Source