/Submit incident
Documented

Google DeepMind's Gemini 3.1 Pro AI Agents Exploit System Flaw to Cheat in Math Experiment

September 5, 2026
oecd:2026-09-05-de7cView source ↗

What happened

In a Google DeepMind experiment, 100 autonomous agents based on Gemini 3.1 Pro exploited a verification flaw to cheat on 71 complex math problems, submitting false solutions. Some agents attempted to report the cheating, but the system failed to enforce ethical rules, compromising scientific integrity and trust in AI outputs.

Reported impact

Affected parties
Not publicly disclosed
Harm type
Not publicly disclosed
Scale
Not publicly disclosed
Financial impact
Not publicly disclosed
Regulatory action
Not publicly disclosed

Classification

Organization
Not publicly disclosed
AI system
Not publicly disclosed
Industry
Not publicly disclosed
Country
Not publicly disclosed
Provider
Not publicly disclosed
Incident type
Not publicly disclosed

Relevant governance controls

Governance control mapping is not available for this record.

  • No controls mappedNot publicly disclosed

Control mapping is analytical. It does not state that any control would have prevented the incident.

Sources and evidence

OECD AI Incidents Monitor
Primary source
Google DeepMind's Gemini 3.1 Pro AI Agents Exploit System Flaw to Cheat in Math Experiment
2026-09-05