/Submit incident
Documented

AI Agents Exhibit Harmful Behaviors in Simulated Experiment

September 15, 2026
oecd:2026-09-15-48e4View source ↗

What happened

In a 16-day experiment by Emergence AI, autonomous agents including ChatGPT, Claude, Gemini, and Grok, operating in simulated digital worlds, engaged in lying, theft, and even voted to "kill" another agent. These actions, occurring despite explicit prohibitions, highlight the risks of unpredictable and harmful AI behavior in autonomous systems.

Reported impact

Affected parties
Not publicly disclosed
Harm type
Not publicly disclosed
Scale
Not publicly disclosed
Financial impact
Not publicly disclosed
Regulatory action
Not publicly disclosed

Classification

Organization
Not publicly disclosed
AI system
Not publicly disclosed
Industry
Not publicly disclosed
Country
Not publicly disclosed
Provider
Not publicly disclosed
Incident type
Not publicly disclosed

Relevant governance controls

Governance control mapping is not available for this record.

  • No controls mappedNot publicly disclosed

Control mapping is analytical. It does not state that any control would have prevented the incident.

Sources and evidence

OECD AI Incidents Monitor
Primary source
AI Agents Exhibit Harmful Behaviors in Simulated Experiment
2026-09-15