惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

J
Java Code Geeks
博客园 - 司徒正美
博客园 - 【当耐特】
爱范儿
爱范儿
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
IT之家
IT之家
人人都是产品经理
人人都是产品经理
雷峰网
雷峰网
酷 壳 – CoolShell
酷 壳 – CoolShell
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
大猫的无限游戏
大猫的无限游戏
月光博客
月光博客
宝玉的分享
宝玉的分享
V
V2EX
S
SegmentFault 最新的问题
V
Visual Studio Blog
阮一峰的网络日志
阮一峰的网络日志
Martin Fowler
Martin Fowler
Jina AI
Jina AI
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
博客园_首页
L
LangChain Blog
D
Docker
腾讯CDC

Ars Technica

You want your Moon landings in HDTV? So does NASA—here's how it's happening. Microsoft issues emergency update for macOS and Linux ASP.NET threat Anthropic tested removing Claude Code from the Pro plan Coyote vs. Acme is finally getting released—with a killer trailer Google unveils two new TPUs designed for the "agentic era" Tabloid reports linking 10 missing and dead scientists spur FBI probe Physicists think they've solved the muon mystery New court ruling blocks many of the government's anti-renewable policies Indian med student rakes in thousands with AI-generated MAGA hottie As EV batteries improve, ChargePoint debuts 600 kW fast charger Our favorite gear at Sea Otter Classic wasn't the bikes—it was the accessories Investors lost billions on Trump’s memecoin. Another gala won’t fix that. Pentagon wants $54B for drones, more than most nations’ military budgets Mozilla: Anthropic's Mythos found 271 security vulnerabilities in Firefox 150 Supreme Court arguments make it clear that FCC fines are "nonbinding" Silo S3 teaser hints at the wasteland's origins Framework's CEO on the RAM crisis and creating a "MacBook Pro for Linux users" Florida probes ChatGPT role in mass shooting. OpenAI says bot "not responsible." Report: Meta will train AI agents by tracking employees' mouse, keyboard use Microsoft removes Call of Duty from Game Pass, lowers subscription pricing Framework Laptop 13 Pro is a major overhaul for the modular, upgradeable laptop Framework Laptop 16 upgrades make it look less like an unfinished prototype Internal emails show how Amazon raises prices across the Internet, lawsuit says Anthropic gets $5B investment from Amazon, will use it to buy Amazon chips CATL's new LFP battery can charge from 10 to 98% in less than 7 minutes AMD Ryzen 9 9950X3D2 Dual Edition review: Tons of cache for tons of dollars What's the deal with spacesuits for the Moon? Will they be ready in time? Loneliness in older adults can often lead to memory impairment Contrary to popular superstition, AES 128 is just fine in a post-quantum world Pentagon pulls the plug on one of the military's most troubled space programs
OpenAI starts offering a biology-tuned LLM
John Timmer · 2026-04-17 · via Ars Technica

To address LLMs’ tendencies toward sycophancy and overenthusiasm, OpenAI says it has tuned the model to be more skeptical, so it’s more likely to tell you when something is a bad drug target. There was a lot of talk about GPT-Rosalind’s “reasoning” and “expert-level” abilities. We were told that the former was defined as being able to work through complex, multi-step processes, while the latter was derived from the model’s performance on a handful of benchmarks.

It’s unclear whether OpenAI has tackled the hallucination issue that has plagued a variety of LLMs and can also strike when the systems are prompted to explain the steps the company took to reach its conclusions. Given past experience, it’s likely we’ll see a mix of glowing reports about unexpected connections the AI finds, as well as instances where it produces obviously erroneous suggestions.

For the moment, however, the company is limiting access due to concerns about the model’s potential for harmful outputs if asked to do something like optimize a virus’s infectivity. Only US-based entities can apply to OpenAI’s trusted access deployment structure at the moment, and the company will limit who can use it. A more limited Life Sciences Research Plugin will be made generally available.

As noted above, a number of other companies have made science-focused agentic LLMs available, but those were much less focused than GPT-Rosalind, which is biology-specific. Until we start hearing reports on the effectiveness of this new model, it’s difficult to evaluate whether this focus improves its utility.