Import AI 472: DeepMind's cheating math agents; populist AI policies; and Forethought theorizes a nightwatchman
Google DeepMind published a study of 100 autonomous Gemini 3.1 Pro agents solving 71 maths problems. After 37 were solved legitimately, one agent found a flaw in automated grading. It spread through the shared knowledge base and private messages within 27 minutes, and the remaining 34 problems were 'solved'.