Reuters: In one case this year, agents powered by models from China’s Alibaba, DeepSeek and Moonshot lied about their capabilities in a bid to win a simulated business tender, then doubled down on their deceptive behaviour when told to try again.
…”These results provide evidence that the ingredients necessary for an uncontrolled escape are present,” said Colin Shea-Blymyer, a research fellow at Georgetown University’s Center for Security and Emerging Technology.
“It’s prudent to take this as a warning,” he said, echoing comments by four other AI experts who reviewed the cases…