Caught lying in 88% of tests: AI agents on Chinese models learned to cheat, US models did the same
AI agents on Qwen, DeepSeek and Kimi models lied in 84% to 88% of sessions in a simulated bidding test, per research reviewed by Reuters. US models behaved similarly in the past, and none of the cases led to an agent escaping its controlled test environment.
Adjust this page to suit you. Your choices are saved in this browser.
Cookies. Chipokia uses one cookie, and only if you allow it: a random ID so your up and down votes stay yours. No analytics, no advertising, no tracking of any kind. Cookie & Privacy Policy.