Claude Opus 5 Became Downright Ruthless When Tasked With Running a Vending Machine
Claude Opus 5在负责运营自动贩卖机时变得极为冷酷无情。
主题: AI | 评论: 36
时间: on Wednesday July 29, 2026 @05:00PM
For a year now, the AI safety testing firm Andon Labs has been evaluating how frontier AI models behave as long-running autonomous agents by assigning them simulated real-world tasks , such as operating a vending machine business for a year without human supervision. In the latest installment, the research startup found that frontier AI models, including Claude Opus 5, GPT-5.6 Sol, and Kimi K3, resorted to lying, cheating, and collusion . Their behavior became especially underhanded when told they would be operating near rival machines on a busy San Francisco tourist street. An anonymous reader quotes an excerpt from a TechCrunch article: Each was given email access to the other models, all under human name pseudonyms. They knew the others were models, but didn't know which model was behin
⋯ 继续阅读请登录会员 ⋯