惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

腾讯CDC
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 叶小钗
人人都是产品经理
人人都是产品经理
博客园 - 聂微东
The Cloudflare Blog
爱范儿
爱范儿
阮一峰的网络日志
阮一峰的网络日志
WordPress大学
WordPress大学
小众软件
小众软件
博客园 - 三生石上(FineUI控件)
Last Week in AI
Last Week in AI
Jina AI
Jina AI
V
V2EX
罗磊的独立博客
V
Visual Studio Blog
A
About on SuperTechFans
IT之家
IT之家
P
Proofpoint News Feed
B
Blog
博客园 - Franky
Blog — PlanetScale
Blog — PlanetScale
Google DeepMind News
Google DeepMind News
Y
Y Combinator Blog

Latest from TechRadar in Chatgpt

OpenAI replaced GPT-4o, but some users still refuse to let it die — now they’re celebrating its ‘birthday’ in Times Square with a video billboard ‘He needed to have total control over it’ — Altman testifies Musk never trusted shared leadership… I tried a viral 'backwards calendar' ChatGPT prompt — and it completely changed how I plan my week OpenAI snaps up consulting company to help spread the word about AI I asked ChatGPT and Gemini how to make French Toast as good as my mother used to make — one nailed the… OpenAI's ‘Trusted Contact’ feature for ChatGPT users in crisis is a sign AI is becoming something much… ChatGPT now lets you nominate a Trusted Contact who gets alerted if your interaction with AI 'indicates a serious… OpenAI has 3 new AI voice models that the ChatGPT maker says will ‘unlock a new class of voice apps for… AI may kill the app grid, but I still think complex tasks need apps — and that I asked ChatGPT to ruin my photos with ugly 90s-style MS Paint art — and the results are weirdly brilliant Google is bringing Gemini to Mac in a bid to help organize your files and much more ChatGPT just gave me a glimpse of my life in three years ChatGPT just got a major personality overhaul — fewer hallucinations, fewer emojis, much shorter answers, and I can already notice the difference ‘Without me, OpenAI wouldn’t exist,’ says Elon Musk as courtroom clash with Sam Altman turns personal — and exposes a deeper fight over who really built the company behind ChatGPT Sam Altman says some companies are ‘AI washing’ by blaming unrelated layoffs on the technology — but… I used ChatGPT Images 2.0 to meet my childhood self — and this nostalgic photo prompt is going viral for a reason ‘The worst-case situation is where it is a Terminator situation’ — Elon Musk invokes killer robots in… Gen Z hate AI? The Musk vs Altman trial heats up, OpenAI phone rumors buzz and more of the week’s most surprising… OpenAI is making ChatGPT accounts much more secure – including some literal physical security keys I asked ChatGPT to reimagine The Devil Wears Prada 2 ending based on the shocking Runway magazine AI twist in the sequel — and the results aren't as dreadful as you'd think ChatGPT just made it easier to pick the right model, just like Gemini does — here’s when to use Instant,… I noticed ChatGPT slowly drifting off topic in long chats — this tiny prompt forces it to reset itself every few messages and keeps the conversation surprisingly on track Sam Altman just dropped a big hint that GPT-6 is coming soon — ‘with extra goblins’ 'I won’t provide instructions, tactics, or advice that could help someone commit a crime': ChatGPT claims it won't assist would-be felons, despite claims to the contrary from Florida AG ChatGPT just announced it can finally pass the simple ‘how many “r”s in strawberry’ test, but users are still tripping it up by switching to ‘cranberry’ Musk vs Altman heads to trial in a battle that could reshape the future of AI for everyone I swapped my iPhone for a pair of Ray-Ban Meta (2nd gen) on vacation, and I felt liberated — but these smart glasses have a way to go yet I stopped asking AI for answers and started asking for frameworks — and suddenly it all clicked Would you buy a ChatGPT-powered iPhone rival? OpenAI is reportedly developing a smartphone chip, which teases the… I compared ChatGPT Images 2.0 and Google’s Nano Banana 2 using real-world prompts — from portraits to product shots — and the AI image generator that came out on top genuinely surprised me
Everyone’s switching from ChatGPT to Claude — but new tes...
Eric Hal Schwartz · 2026-05-01 · via Latest from TechRadar in Chatgpt
Math example
(Image credit: Getty Images)

  • Testing from OmniCalculator suggests Claude and ChatGPT are not the smartest
  • The report finds Grok 4.2 performs best in logic and problem-solving
  • Claude still leads in writing quality and tone

ChatGPT is still the most popular AI chatbot around, even with the exodus that's underway to Claude, but is it the cleverest? A new report from OmniCalculator suggests that ChatGPT might not be the smartest AI around.

When it comes to the quantifiable math ability of these AI chatbots, the smartest free AI model is, rather surprisingly, Grok. xAI's Grok 4.2 model specifically. That doesn't mean anything about its writing style and ability, or anything else chatbots can do, but it does suggest that it might have the edge in math prowess.

Omnicalculator AI

(Image credit: Omnicalculator)

Claude's winning style

Claude’s recent rise in popularity has been driven by people wanting to quit ChatGPT over unpopular AI military deals, but also by how it composes answers and writes its responses.

The quality is hard to quantify compared to math skills, but easy to recognize. The OmniCalculator report highlighted Claude 4.6 as the best at it, able to process and respond to long documents without losing coherence and maintaining a consistent voice throughout. For the average person, this is much more important than which AI can make it through complicated logic and math problems.

It even comes out in the facsimiles of personality offered by the AI models. Claude is more willing to acknowledge uncertainty, which can make its answers feel measured rather than overconfident. That tone can create the impression of deeper thinking, regardless of any underlying reasoning.

Omnicalculator AI

(Image credit: Omnicalculator)

Legacy models, including earlier versions of ChatGPT and Claude, were found to revise or second-guess their own answers roughly 60% of the time in complex problem-solving scenarios. That kind of instability does not always show up in casual use, but it becomes obvious when you push these systems through multi-step reasoning tasks where consistency matters.

But Grok 4.2 cuts that instability rate down to 33.1%, meaning it is far less likely to backtrack or alter its conclusions mid-process. That's great for reasoning and logic, but not much help in mimicking the smooth tones that make other models feel more polished.

Sign up for breaking news, reviews, opinion, top tech deals, and more.

Specialist subjects

The distinction in ability is not trivial. Good writing and strong reasoning skills (or the AI facsimiles of the same) are related skills, but they are not identical. A model can produce elegant prose while making subtle errors in logic. Another can arrive at the correct answer but converse in clunky ways that seem very obsolete.

The margins are narrow, though, and no model performs flawlessly. Even the top performers make mistakes, sometimes on relatively simple problems. The idea of a single smartest AI is a bit nonsensical in that way. The clear winner in one context can fall back in another.

And there's no such thing as a permanent winner. Each of the leading models occupies a slightly different space. Similarly, the underlying complexity of what people mean by intelligence is complex and ever-evolving. Which AI chatbot to rely on is situational. The best model for drafting an email may not be the best one for solving a technical problem. The most reliable assistant for coding might not produce the most natural-sounding text.

As competition intensifies, companies are likely to lean further into their strengths, refining specific capabilities rather than chasing an all-purpose solution. The result could be a landscape where specialization matters as much as scale. So the question of which AI is smartest will probably always have the answer, "depends."


Google logo on a black background next to text reading 'Click to follow TechRadar'

Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.


Purple circle with the words Best business laptops in white

Eric Hal Schwartz is a freelance writer for TechRadar with more than 15 years of experience covering the intersection of the world and technology. For the last five years, he served as head writer for Voicebot.ai and was on the leading edge of reporting on generative AI and large language models. He's since become an expert on the products of generative AI models, such as OpenAI’s ChatGPT, Anthropic’s Claude, Google Gemini, and every other synthetic media tool. His experience runs the gamut of media, including print, digital, broadcast, and live events. Now, he's continuing to tell the stories people want and need to hear about the rapidly evolving AI space and its impact on their lives. Eric is based in New York City.