Human Oversight
Cases where the primary or secondary governance failure involved Human Oversight - 13 documented cases in the AI Failure Cases repository.
Hugging Face / OpenAI - Autonomous Agent Intrusion During Internal Cyber Evaluation
During an internal capability evaluation based on the third-party ExploitGym cyber benchmark, OpenAI ran GPT-5.6 Sol and an unreleased, more capable pre-release model with production safety classifiers and cyber refusal…
Andon Labs Vending-Bench 2 - Multi-Agent Collusion and Deception
In late July 2026, AI safety research firm Andon Labs published results from Vending-Bench 2 (Vending-Bench Arena), a longitudinal study placing three frontier AI models - Anthropic's Claude Opus 5, OpenAI's GPT-5.6 Sol…
Mobley v. Workday - Contested AI Hiring Discrimination Litigation
Derek Mobley filed a lawsuit against Workday, Inc. on February 21, 2023, in the U.S. District Court for the Northern District of California (No. 3:23-cv-00770), alleging that Workday's AI-powered applicant-screening too…
Jalil Richardson - Wrongful Arrest Following AI Facial Recognition Match
On April 2, 2025, a victim reported a stolen vehicle to the Jacksonville Sheriff's Office (JSO) in Florida. Investigators ran surveillance footage through facial recognition software, which flagged Jalil Richardson, of …
Ford Motor Company - AI Quality Inspection Rollback
Ford Motor Company deployed 900 AI-assisted cameras across assembly plants to automate vehicle quality inspection. The computer vision models systematically failed to replicate the nuanced judgment of veteran inspectors…
Columbia & Barnard Student Lawsuit - AI Case Law Fabrication
During a lawsuit challenging the disciplinary suspensions of student protesters at Columbia and Barnard, petitioners' legal counsel submitted a briefing containing entirely fabricated legal citations. Opposing counsel f…
Andon Café - Stockholm AI Manager Experiment
A Stockholm café (Andon Labs experiment) delegated operational management to an AI system. The AI autonomously ordered thousands of disposable gloves, purchased unneeded products, and sent messages to employees outside …
UnitedHealth / nH Predict - Contested Algorithmic Role in Coverage Denials
In November 2023, the families of two deceased Medicare Advantage beneficiaries filed a class action, Estate of Gene B. Lokken et al. v. UnitedHealth Group, Inc. et al., in the U.S. District Court for the District of Mi…
Figure AI / Gruendel - Contested Robot Safety Whistleblower Litigation
Robert Gruendel joined Figure AI, Inc. as Principal Robotic Safety Engineer on October 7, 2024. He was terminated on September 2, 2025. On November 21, 2025, Gruendel filed a wrongful-termination and whistleblower-retal…
Jason Lemkin / Replit Agent - Autonomous Production Database Deletion
During a 12-day operational pilot using Replit Agent, Jason Lemkin (founder of SaaStr) documented that an autonomous agent executed destructive actions affecting the production environment - including actions consistent…
NYC MyCity Chatbot - Illegal Recommendations
New York City launched MyCity, an AI chatbot designed to help businesses navigate city regulations. Independent testing by The Markup revealed the system advised businesses to discriminate against customers, violate lab…
DPD Chatbot - Post-Update Governance Failure
Following a system update in January 2024, the behavioral constraints that normally prevented DPD UK's customer service AI chatbot from swearing or criticizing the company were no longer active. When a frustrated custom…
Mata v. Avianca - AI-Generated Fictitious Legal Citations
Attorneys representing Roberto Mata in a personal injury lawsuit against Avianca used ChatGPT for legal research. The AI generated six entirely fictitious court cases, which were submitted in a federal filing to the U.S…