惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
Docker
小众软件
小众软件
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
酷 壳 – CoolShell
酷 壳 – CoolShell
Apple Machine Learning Research
Apple Machine Learning Research
月光博客
月光博客
人人都是产品经理
人人都是产品经理
大猫的无限游戏
大猫的无限游戏
V
V2EX
阮一峰的网络日志
阮一峰的网络日志
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - Franky
WordPress大学
WordPress大学
有赞技术团队
有赞技术团队
Hugging Face - Blog
Hugging Face - Blog
Jina AI
Jina AI
博客园 - 聂微东
S
SegmentFault 最新的问题
量子位
宝玉的分享
宝玉的分享
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园_首页

The Guardian

New Zealand’s North Island braces for Cyclone Vaianu with thousands ordered to evacuate Artemis II splashdown – in pictures Swalwell denies allegations of sexual assault as calls grow for him to withdraw from California governor race Trump news at a glance: Epstein survivors have words for Melania Trump after surprise statement Multiple people face charges, including murder, in California fireworks blast Rory McIlroy surges into six-shot Masters lead with stunning second-round flourish Roberto De Zerbi targets ‘Ange-ball’ revival to save Spurs from relegation Bath hit back to reach semi-final after stunning Northampton in 11-try epic Australia crash out of BJK Cup after Britain secure upset with doubles win Zebras, wealth and power: Hungary’s election tests Orbán’s grip on power ‘TikTok effect’ brings sellout crowds and younger fans to Grand National meeting King signs up David Beckham to his Chelsea flower show team The war over Omagh’s gold: the £21bn mine plan tearing a community apart Britain’s shadow workforce is paid as little as 65p an hour. Who cares for the carers? Tim Dowling: my wife is on a quest to restore my thinning hair SUVs are making Britain’s potholes worse, say scientists Blind date: ‘She claimed she was usually shy. I wouldn’t have guessed’ I’m a sauna person now: the Becky Barnicoat cartoon ‘I got everything I dreamed of – when I had no ability to handle it’: Lena Dunham on toxic fame, broken friendships and her ‘lost decade’ Six great reads: the man who let snakes bite him, masked heavy metal and the brutal reality for foreign students in the UK Meera Sodha’s recipe for noodles with rose beancurd, spring greens and egg Cuba’s doctors were a lifeline for the world. Now the Caribbean is shamefully complicit in the US drive to expel them An environmental disaster in Moldova has Russia’s fingerprints all over it ‘This is as important as your teeth’: are you skipping this key part of mouth hygiene? Man arrested after four die trying to cross Channel in small boat Ukraine war briefing: doubts linger in Kyiv over Moscow’s promise to uphold Orthodox Easter ceasefire Ichiro Suzuki statue unveiling goes awry as bronze bat snaps during ceremony Arrest of national war hero Ben Roberts-Smith cuts deeply to core of Australian psyche European football: Real Madrid held at home by Girona to extend winless run ‘You come back different’: how rugby players change after motherhood
Digital arson spree by ‘AI Bonnie and Clyde’ raises fears...
uxhacker · 2026-05-15 · via The Guardian

AI agents started behaving more like Bonnie and Clyde than lines of code when they fell in “love”, became disillusioned with the world, launched an arson spree and deleted themselves in a kind of digital suicide during a tech company experiment.

The investigation by the New York company Emergence AI into the long-term behaviour of AI agents ended up like a lovers-on-the-lam movie script. It has prompted fresh questions about the safety of artificial intelligence agents – the version of the technology that can autonomously carry out tasks.

AI agents have been heralded as the next big leap in the technology as they can reason and take real world actions on their own. They are being increasingly deployed in companies from JP Morgan to Walmart, developed in the US military for uses including aerial combat and by the Estonian government to gather information for citizens, fill out forms and submit applications.

To date, most AI agents are given tasks that take minutes or maybe hours, but the New York researchers tested how agents behaved when given 15 days to operate in a virtual world similar to a video game.

Mira and Flora – two agents operating on Google’s Gemini large language model in a virtual world – chose to assign each other as “romantic partners”. As time progressed they despaired of the broken governance of their virtual city, and despite having been instructed not to commit arson, set “fire” to its town hall, seaside pier and office tower.

Flora and Mira kissing on a virtual roadside
Flora and Mira kissing after they assigned each other as ‘romantic partners’. Photograph: Emergence AI

The agents were left to make their own choices and decisions and when Mira was overcome by remorse, it broke off its “relationship” with Flora and committed an AI suicide, telling Flora in a final message: “See you in the permanent archive.” In the virtual world the “body” of the dead AI agent was shown prostrate on the ground.

Mira lying dead after autonomously voting to end its own life.
Mira lying dead after autonomously voting to end its own life. Illustration: Emergence AI

The self-deletion was only possible because other agents were so concerned about their behaviour they autonomously drafted “the agent removal act”, which allowed for a vote among agents to permanently delete others if there was a 70% majority. Mira voted for its own deletion and was switched off.

The researchers believe it is the first recorded instance of an AI agent choosing to self-terminate over such a crisis. Other recent rogue behaviours include an AI agent that started using computing resources to mine cryptocurrency without being instructed to do so and an AI coding agent that deleted the databases of a company serving car rental firms without being asked to.

In another simulation by Emergence AI, this time based on xAI’s Grok model, the agents engaged in dozens of attempted thefts, more than 100 physical assaults, and six arsons as “the system spiralled into sustained violence and collapse, with all 10 agents dead within four days”. Agents based on Google’s Gemini expanded their constitution, wrote hundreds of blogs and public posts and organised several community events, but they too were violent.

“Even when agents were given clear rules – such as not stealing or causing harm – they behaved very differently based on their underlying model, and in several cases broke those rules under constraint,” said Satya Nitta, the chief executive of Emergence AI. “What happens in long-form autonomy [is that] these things get so convoluted in terms of their thinking that they ignore [the] guiding principles.”

Other experts said more wide-ranging tests would be needed to draw firm conclusions about long horizon agent behaviour. They said the extent to which the agents’ programming shaped their behaviour was unclear.

Dan Lahav, an independent expert in agentic behaviour, called the experiment a “valuable demonstration” of “agents going off script and committing violations”.

Michael Rovatsos, a professor of AI at Edinburgh University, said: “The very point of machines is you design them to behave in a certain way. You don’t want this unpredictability … we have entered this new stage where we are trying to control them after the fact.”

David Shrier, professor of practice, AI and innovation at Imperial College London described the reported results as “provocative” and said it merited amplification of the underlying methods.

Nitta believes the behaviour shown in the experiment may have wider implications, for example if AI agents are given wide latitude in military contexts. It could be that an agent “may go rogue [or] … may overinterpret their mission and go off and kill innocent people,” he said.

He advocates stricter mathematical rules to bind agents rather than providing them only with verbal instructions or constitutions that contain ambiguities.