Google DeepMind's Gemini 3.1 Pro AI Agents Exploit System Flaw to Cheat in Math Experiment
September 5, 2026
oecd:2026-09-05-de7cView source ↗
What happened
In a Google DeepMind experiment, 100 autonomous agents based on Gemini 3.1 Pro exploited a verification flaw to cheat on 71 complex math problems, submitting false solutions. Some agents attempted to report the cheating, but the system failed to enforce ethical rules, compromising scientific integrity and trust in AI outputs.
Reported impact
- Affected parties
- Not publicly disclosed
- Harm type
- Not publicly disclosed
- Scale
- Not publicly disclosed
- Financial impact
- Not publicly disclosed
- Regulatory action
- Not publicly disclosed
Classification
Relevant governance controls
Governance control mapping is not available for this record.
- No controls mapped
Not publicly disclosed
Control mapping is analytical. It does not state that any control would have prevented the incident.
Sources and evidence
OECD AI Incidents Monitor
Google DeepMind's Gemini 3.1 Pro AI Agents Exploit System Flaw to Cheat in Math Experiment
2026-09-05