惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

F
Fortinet All Blogs
有赞技术团队
有赞技术团队
量子位
N
Netflix TechBlog - Medium
博客园 - 叶小钗
博客园 - 三生石上(FineUI控件)
Google DeepMind News
Google DeepMind News
aimingoo的专栏
aimingoo的专栏
GbyAI
GbyAI
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Blog — PlanetScale
Blog — PlanetScale
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
月光博客
月光博客
Martin Fowler
Martin Fowler
Y
Y Combinator Blog
宝玉的分享
宝玉的分享
博客园 - 司徒正美
云风的 BLOG
云风的 BLOG
V
Visual Studio Blog
V
V2EX
IT之家
IT之家
L
LangChain Blog
大猫的无限游戏
大猫的无限游戏
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More

SiliconANGLE

Will agentic AI governance run amok? The lesson of Asimov’s Three Laws - SiliconANGLE AI + quantum, Amazon vs. Starlink and the wide-open US-China internet battle - SiliconANGLE Team Cymru launches Total Insights Feed to replace legacy threat intelligence lists - SiliconANGLE AI Mode in Chrome adds split-screen view to enhance the web search experience - SiliconANGLE Resolve AI raises $40M at $1.5B valuation to optimize production environments - SiliconANGLE How Zscaler and OpenAI turn zero-trust security into an AI accelerator - SiliconANGLE OpenAI ratchets up Codex's agentic capabilities to rival Claude Code - SiliconANGLE Anthropic launches Claude Opus 4.7 with coding, visual reasoning improvements - SiliconANGLE Slash raises $100M at a $1.4B valuation to expand AI-powered banking platform for online businesses - SiliconANGLE Canva unveils Canva AI 2.0, recasting its platform as an agentic system for work - SiliconANGLE Data center, consumer device chips boost TSMC’s revenue - SiliconANGLE Mission-critical security cannot be bolted on, says Oracle - SiliconANGLE Agentic infrastructure reshapes enterprise AI - SiliconANGLE Data quality, and data freedom, foundational for AI success - SiliconANGLE Data trust is a bedrock in successful, scalable AI outcomes - SiliconANGLE Google introduces new agentic AI-ready tools and resources for Android developers  - SiliconANGLE Agentic AI orchestration separates winners from laggards - SiliconANGLE Data-driven tools turning the tide against human trafficking - SiliconANGLE Achieving trusted AI development goes beyond 'vibes' - SiliconANGLE Impinj boosts edge computing power in updated R700 RAIN RFID reader - SiliconANGLE Certinia powers professional services with AI - SiliconANGLE Antioch prepares to accelerate simulated testing for autonomous robots after raising $8.5M - SiliconANGLE Developer tooling startup Expo nabs $45M investment - SiliconANGLE Solidroad lands $25M to bring AI to customer support interactions - SiliconANGLE DuploCloud lands compliance and AI governance certifications as enterprise buyers tighten scrutiny - SiliconANGLE Lua lands $5.8M to help businesses build and manage AI agent workforces - SiliconANGLE Best of frenemies: Oracle's and AWS' clouds unite with dedicated, private connectivity - SiliconANGLE NIST shifts National Vulnerability Database to risk-based triage as CVE submissions hit record levels - SiliconANGLE Cisco goes to the races with new Churchill Downs multiyear partnership - SiliconANGLE Susecon 2026 will tackle the future of open-source platforms - SiliconANGLE
OpenAI releases GPT-5.5 with advanced math, coding capabi...
Maria Deutsc · 2026-04-24 · via SiliconANGLE

OpenAI releases GPT-5.5 with advanced math, coding capabilities

OpenAI Group PBC today launched a new large language model that is significantly better than its predecessors at solving math problems and writing code.

GPT-5.5 is rolling out a week after rival Anthropic PBC released its latest LLM. OpenAI is offering the model in two flavors: a standard version and a more capable, significantly pricier edition called GPT-5.5 Pro.

The company says that both variants deliver output quality improvements in multiple areas. The standard edition of GPT-5.5 is more adept than its predecessor at computer use tasks and knowledge work. GPT-5.5 Pro, in turn, provides particularly large quality gains across business, legal, education and data science use cases.

GPT-5.5 is also better at interpreting ambiguous instructions. Historically, LLM users had to describe each step of the task they sought to automate or risk output errors. By contrast, GPT-5.5 can automatically figure out details such as how to use an MCP server even if the user doesn’t provide an explanation.

OpenAI compared GPT-5.5 to Claude Opus 4.7, the new LLM that Anthropic debuted last week, across more than a dozen benchmarks. The standard and Pro editions of the former model performed better across many of the tests.

One of the most difficult benchmarks in OpenAI’s test suite is FrontierMath Tier 4. It comprises dozens of postdoctoral-level math problems that can take a human expert upwards of days to solve. GPT 5.5 Pro scored 39.6%, nearly double the 22.9% that Claude Opus 4.7 achieved.

OpenAI says that a customized version of GPT-5.5 helped researchers discover a new proof, a series of equations that confirms a mathematical theorem. The proof related to objects known as Ramsey numbers. Such objects are a major focus of a mathematical field called combinatorics that has broad computer science applications.

According to OpenAI, GPT-5.5 is also better than competing models at many programming tasks. The standard version of the LLM achieved a 82.7% score on Terminal-Bench 2.0, which measures LLMs’ ability to use command line tools. Claude Opus 4.7 scored 69.4%.

OpenAI says that it has already put GPT-5.5’s coding skills to use internally. The LLM helped optimize the software that manages the infrastructure on which it runs. That hardware comprises Nvidia Corp.’s GB200 and GB300 NVL72 systems, which include the chipmaker’s Blackwell B200 and Blackwell Ultra graphics processing units, respectively. 

GPUs have significantly more cores than a central processing unit. OpenAI’s infrastructure management software groups the LLM requests that are sent to a GPU into batches, or chunks, and distributes them across the chip’s cores. According to the company, GPT-5.5 developed a more efficient way of going about the task that increased token generation speeds by over 20%.

The model is also proficient at less technical tasks. It set a record on GDPval, a benchmark dataset that tests LLMs’ ability to complete economically valuable tasks across 44 fields. Notably, the standard version of GPT-5.5 bested both the Pro edition and Claude Opus 4.7 with a 84.9% score. 

GPT 5.5 is available in ChatGPT and Codex for users with Plus, Pro, Business, and Enterprise subscriptions. GPT-5.5 Pro is only available in the latter 3 plans through ChatGPT. OpenAI will bring the LLM to its application programming interface “very soon.”

Image: OpenAI

A message from John Furrier, co-founder of SiliconANGLE:

Support our mission to keep content open and free by engaging with theCUBE community. Join theCUBE’s Alumni Trust Network, where technology leaders connect, share intelligence and create opportunities.

  • 15M+ viewers of theCUBE videos, powering conversations across AI, cloud, cybersecurity and more
  • 11.4k+ theCUBE alumni — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network.

About SiliconANGLE Media

SiliconANGLE Media is a recognized leader in digital media innovation, uniting breakthrough technology, strategic insights and real-time audience engagement. As the parent company of SiliconANGLE, theCUBE Network, theCUBE Research, CUBE365, theCUBE AI and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.

Founded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.