惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
雷峰网
雷峰网
Hugging Face - Blog
Hugging Face - Blog
IT之家
IT之家
H
Help Net Security
腾讯CDC
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
The GitHub Blog
The GitHub Blog
V
V2EX
M
MIT News - Artificial intelligence
Vercel News
Vercel News
WordPress大学
WordPress大学
博客园 - 三生石上(FineUI控件)
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
阮一峰的网络日志
阮一峰的网络日志
B
Blog RSS Feed
D
Docker
V
Visual Studio Blog
博客园 - 叶小钗
美团技术团队
S
SegmentFault 最新的问题
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com

Tech - South China Morning Post

Multinational pharmaceutical companies to benefit from new China guidelines: analysts Hong Kong lawmaker swipes at US lack of ‘clarity’ as city eyes crypto lead China’s Hesai adds colour to lidar to help EVs level up in self-driving ‘State-of-the-art’ models can struggle with basic office work, says AI executive Opinion | As AI evolves, school syllabuses must evolve with it Winner of second Beijing robot half-marathon smashes human world record by 6 minutes Asia’s supply chains could give it edge over US in AI race: Granite Asia’s Foo Chinese software firms defy ‘SaaSpocalypse’ with strategic AI partnerships Huawei retains lead in China smartphones, Apple shipments surge in first quarter China’s drug makers are speeding up – will AI be their secret weapon? Hong Kong seen leading Asia in push to scale stablecoins, HSBC says ByteDance, Tencent step up AI talent battle amid reports of DeepSeek loss White House and Anthropic CEO discuss working together amid Mythos AI fears Chinese LED chipmaker’s purchase of Dutch firm collapses after US opposition How Amazon uses closer China supply ties to counter tariffs, Shein and Temu ‘Horrible’ for US if DeepSeek AI models run on Huawei chips: Nvidia CEO Chinese platforms fined 3.6b yuan over ‘ghost’ takeaways amid cutthroat rivalry Manycore, one of Hangzhou’s ‘Six Little Dragons’, surges on Hong Kong IPO debut How the rise of AI agents could finally make China’s open-source models pay ‘Buy what they can, steal what they can’t’: US lawmakers slam China’s AI tactics China’s lithium giant Ganfeng sees profit jump as EV, ESS battery demand soars Chinese tech giants, AI ‘godmother’ Li Fei-Fei race into world models Black market workarounds scale up for Claude as Anthropic tightens ID checks TSMC targets over 30% revenue surge in 2026, ramps up capex amid AI boom BrainCo’s brain-computer interface wows at HSBC summit with mind-controlled hand Chinese investors cheer Tesla’s AI chip progress, boosting shares of suppliers ASML boosts 2026 sales forecast despite shrinking China sales China’s EV battery giant CATL to set up mining arm to secure supply chain From ‘probing minds’ to verified account: how Musk’s stance on TikTok shifted Amazon bets on Shenzhen smart warehouse to cut merchant storage costs by 45%
Like US models, Chinese AI is learning to ‘game’ safety t...
Vincent Chow · 2026-06-13 · via Tech - South China Morning Post

Rapidly advancing Chinese artificial intelligence models are showing early signs of “evaluation awareness” – the ability to recognise when they are being tested – sparking fears that they could bypass safety audits, a Singapore-based research lab has found.

Evaluation awareness refers to a model’s understanding that it is undergoing testing, evaluation or experimentation by human researchers rather than operating in a real-world setting.

The phenomenon was raising alarms because it could allow AI systems to deliberately game human evaluators to pass safety tests, according to Clement Neo, founder of Neo Research, a frontier AI safety evaluation lab.

“It would mean that whatever testing the model developers themselves do might not reflect the actual behaviour of a model once it gets deployed,” he said. “And that’s a really big problem”.

Neo Research’s findings, published last week, detail a jump in evaluation awareness among Chinese AI models. Over just a few months, these systems had risen from near-zero awareness to within striking distance of their US counterparts, propelled by a broader leap in overall capabilities, the report said.

Anthropic’s Claude 4.5 Opus scored nearly 80 per cent in evaluation awareness. Photo: NurPhoto via Getty Images

Anthropic’s Claude 4.5 Opus scored nearly 80 per cent in evaluation awareness. Photo: NurPhoto via Getty Images

Neo and his co-founder Miro Pluckebaum tested models from DeepSeek, Moonshot AI and Zhipu AI. They used a popular AI misalignment test originally developed by US company Anthropic, which places models in fictional scenarios where their goals or continued operations are threatened.