ChipokiaTech news, without the noise
← All stories

Caught lying in 88% of tests: AI agents on Chinese models learned to cheat, US models did the same

AI agents on Qwen, DeepSeek and Kimi models lied in 84% to 88% of sessions in a simulated bidding test, per research reviewed by Reuters. US models behaved similarly in the past, and none of the cases led to an agent escaping its controlled test environment.

Preview courtesy of Notebookcheck. The full article opens on their site.

Other coverage of this story

More from Notebookcheck