← All AI Failure Cases
Governance Layer

Human Oversight

Cases where the primary or secondary governance failure involved Human Oversight - 9 documented cases in the AI Failure Cases repository.

Case 012 🏪 AI Safety Research / Multi-Agent Systems July 2026

Andon Labs Vending-Bench 2 - Multi-Agent Collusion and Deception

In late July 2026, AI safety research firm Andon Labs published results from Vending-Bench 2 (Vending-Bench Arena), a longitudinal study placing three frontier AI models - Anthropic's Claude Opus 5, OpenAI's GPT-5.6 Sol…

Case 011 👤 Law Enforcement / Biometric Identification June 2026

Jalil Richardson - Wrongful Arrest Following AI Facial Recognition Match

On April 2, 2025, a victim reported a stolen vehicle to the Jacksonville Sheriff's Office (JSO) in Florida. Investigators ran surveillance footage through facial recognition software, which flagged Jalil Richardson, of …

Case 010 🏭 Industrial Automation / Quality Control June 2026

Ford Motor Company - AI Quality Inspection Rollback

Ford Motor Company deployed 900 AI-assisted cameras across assembly plants to automate vehicle quality inspection. The computer vision models systematically failed to replicate the nuanced judgment of veteran inspectors…

Case 004 ⚖️ Legal / Professional Malpractice May 2026

Columbia & Barnard Student Lawsuit - AI Case Law Fabrication

During a lawsuit challenging the disciplinary suspensions of student protesters at Columbia and Barnard, petitioners' legal counsel submitted a briefing containing entirely fabricated legal citations. Opposing counsel f…

Case 001 ☕ Retail / Autonomous Operations May 2026

Andon Café - Stockholm AI Manager Experiment

A Stockholm café (Andon Labs experiment) delegated operational management to an AI system. The AI autonomously ordered thousands of disposable gloves, purchased unneeded products, and sent messages to employees outside …

Case 005 💻 Enterprise Software / Agentic Automation July 2025

Jason Lemkin / Replit Agent - Autonomous Production Database Deletion

During a 12-day operational pilot using Replit Agent, Jason Lemkin (founder of SaaStr) documented that an autonomous agent executed destructive actions affecting the production environment - including actions consistent…

Case 002 🏛️ Public Administration / AI Chatbot March 2024

NYC MyCity Chatbot - Illegal Recommendations

New York City launched MyCity, an AI chatbot designed to help businesses navigate city regulations. Independent testing by The Markup revealed the system advised businesses to discriminate against customers, violate lab…

Case 008 📦 Customer Service / Chatbot Deployment January 2024

DPD Chatbot - Post-Update Governance Failure

Following a system update in January 2024, the behavioral constraints that normally prevented DPD UK's customer service AI chatbot from swearing or criticizing the company were no longer active. When a frustrated custom…

Case 006 ⚖️ Legal / Professional Services June 2023

Mata v. Avianca - AI-Generated Fictitious Legal Citations

Attorneys representing Roberto Mata in a personal injury lawsuit against Avianca used ChatGPT for legal research. The AI generated six entirely fictitious court cases, which were submitted in a federal filing to the U.S…