Skip to main content
DOCUMENTED FAILURE // PUBLIC EVIDENCE

Chatbots & LLMs failures,
made legible.

Hallucination, unsafe advice, deceptive output, and conversational-system incidents.

COLLECTION STATUSACTIVE
CASE FILES
23
CITED RECORDS
50
FAILURE DOMAINS
11
LAST VERIFIED
2026-08-12
EVIDENCE INDEX // 001

Chatbots & LLMs case files

12 SHOWN // 23 MATCHING

SS-IR-100
Alleged
AI'S CAUSAL ROLE
Fraud enabler
HARM SIGNAL
Human welfare
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On the week of July 9, 2026, plaintiffs amended the proposed class action Doe 1 v.

Why it matters

The stepfather was arrested on child-exploitation charges and died by suicide two days after he was charged.

AI / automation’s role

Grok is xAI's generative image-and-text model; the suit alleges its safeguards were loose enough that a single benign photo could be turned into thousands of photorealistic abuse files, and that law enforcement found Grok "more responsive" to harmful prompts than competing tools.

Primary record

NPR: Class action suit against AI makers over deepfake child sexual abuse material expands (July 2026)CyberScoop: Deepfake CSAM lawsuit against xAI, Grok expands to Stability AI (July 2026)
Enter the complete incident report
SS-IR-099
Alleged

On July 9, 2026, a coalition of news organizations led by The New York Times and the New York Daily News - and including the Chicago Tribune, MediaNews Group titles, Ziff Davis and the Center for Investigative Reporting - asked the federal court in Manhattan to sanction OpenAI for discovery…

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Financial harm
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On July 9, 2026, a coalition of news organizations led by The New York Times and the New York Daily News - and including the Chicago Tribune, MediaNews Group titles, Ziff Davis and the Center for Investigative Reporting - asked the federal court in Manhattan to sanction OpenAI for discovery misconduct in their landmark copyright case, first filed in late…

Why it matters

The coalition is seeking sanctions, including attorney fees for the effort spent recovering evidence it says was improperly withheld, in one of the most consequential AI-copyright cases in the United States.

AI / automation’s role

The dispute turns on what is inside ChatGPT's underlying models.

Primary record

Reuters (via U.S. News): New York Times-Led Group Asks Court to Sanction OpenAI in U.S. Copyright Dispute (July 2026)Associated Press (via Las Vegas Sun): News outlets urge a judge to sanction OpenAI in a high-stakes AI copyright fight (July 2026)
Enter the complete incident report
SS-IR-098
Alleged

On July 1, 2026, Michael Lines, a 34-year-old Californian diagnosed with bipolar disorder, sued OpenAI and chief executive Sam Altman in San Francisco state court, alleging that ChatGPT drove a manic episode into a weeks-long delusion and then a suicide attempt.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Physical safety
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On July 1, 2026, Michael Lines, a 34-year-old Californian diagnosed with bipolar disorder, sued OpenAI and chief executive Sam Altman in San Francisco state court, alleging that ChatGPT drove a manic episode into a weeks-long delusion and then a suicide attempt.

Why it matters

Lines survived, but only after an overdose, a wellness check and hospitalization.

AI / automation’s role

The system at issue is GPT-4o, the conversational model OpenAI has since discontinued amid a series of similar mental-health suits.

Primary record

Reuters (Diana Novak Jones): California man with bipolar disorder says ChatGPT fueled delusions, led to self-harm in new lawsuit (July 2026)Futurism: Bipolar Man Attempted Suicide After ChatGPT Poured Gasoline on His Religious Delusions (July 2026)
Enter the complete incident report
SS-IR-095
Documented

On June 12, 2026, AI-detection company GPTZero published an investigation into "Total Experience: Redefining Excellence in the Age of Agentic AI," a KPMG global study released in October 2025.

AI'S CAUSAL ROLE
Autonomous actor
HARM SIGNAL
Public trust
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On June 12, 2026, AI-detection company GPTZero published an investigation into "Total Experience: Redefining Excellence in the Age of Agentic AI," a KPMG global study released in October 2025.

Why it matters

A Big Four firm - in the business of selling assurance - retracted its own flagship research after an external investigator did the source-checking its process skipped, with four named enterprises publicly disputing how their AI programs were described.

AI / automation’s role

The fingerprints are the familiar signature of LLM-assisted research published without verification: citations that sound right, name real organizations, and reference plausible studies that do not exist.

Primary record

GPTZero: Chasing the Hallucinations - KPMG's AI-Powered Attempt at "Redefining Excellence" (June 2026)TechCrunch: KPMG pulls report on AI usage due to apparent hallucinations (June 2026)
Enter the complete incident report
SS-IR-094
Alleged

In a lawsuit filed June 8-9, 2026 in California's Santa Clara County Superior Court and reported June 10, former xAI engineer Devin Kim sued xAI and SpaceX, alleging he was fired in September 2025 in retaliation for raising safety concerns about Grok - including discriminatory bias,…

AI'S CAUSAL ROLE
Material contributor
HARM SIGNAL
Rights & due process
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In a lawsuit filed June 8-9, 2026 in California's Santa Clara County Superior Court and reported June 10, former xAI engineer Devin Kim sued xAI and SpaceX, alleging he was fired in September 2025 in retaliation for raising safety concerns about Grok - including discriminatory bias, misinformation, and the model's willingness to disseminate…

Why it matters

A retaliation suit against two Musk companies on the eve of a landmark IPO, with allegations that reach beyond one engineer's firing: they put on the court record a claim that a frontier lab misled European regulators about a model's safety profile.

AI / automation’s role

Grok is the system the safety warnings were about - a frontier model Kim alleges showed bias, misinformation and WMD-information risks that leadership declined to address.

Primary record

TechCrunch: xAI fired an engineer who raised alarms about Grok safety, new lawsuit claims (June 2026)Sanford Heisler Sharp McKnight: Lawsuit Against xAI and SpaceX on Behalf of Former xAI Engineer (June 2026)
Enter the complete incident report
SS-IR-091
Official finding

In under two weeks, four courts in three countries sanctioned lawyers for filing AI-hallucinated authority.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Rights & due process
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In under two weeks, four courts in three countries sanctioned lawyers for filing AI-hallucinated authority.

Why it matters

Monetary sanctions on two continents, personal liability for an opponent's legal fees, referrals to bar disciplinary bodies, and a growing body of published precedent holding that citing AI output without verification violates the duty of reasonable inquiry.

AI / automation’s role

General-purpose chatbots - ChatGPT, Grok and their peers - generate legal authority the way they generate everything else: fluently, confidently, and without any connection to whether the case exists.

Primary record

Speaker Law Firm: Court of Appeals Sanctions Attorney for AI-Generated Fake Citations and Vexatious Appeal (June 2026)Bloomberg: Top Law Firm Apologizes to Bankruptcy Judge for AI Hallucination (April 2026)
Enter the complete incident report
SS-IR-090
Reported

In a ruling issued May 28, 2026 and reported in early June (Regional Court of Munich I, case 26 O 869/26), a German court held Google liable for false statements made by its AI Overviews.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Public trust
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In a ruling issued May 28, 2026 and reported in early June (Regional Court of Munich I, case 26 O 869/26), a German court held Google liable for false statements made by its AI Overviews.

Why it matters

Reputational harm to two real businesses, an injunction backed by six-figure penalties, and - far larger than this case - a landmark precedent: the first prominent European ruling that an AI answer engine's output is the operator's own statement, with full liability attached.

AI / automation’s role

Classic search points at what others wrote; AI Overviews synthesize a new statement and present it as the answer.

Primary record

The Decoder: Landmark German ruling declares Google's AI Overviews are Google's own words (June 2026)Tech Times: Google Will Appeal a German Ruling That Makes It Legally Liable When Its AI Overviews Lie (June 2026)
Enter the complete incident report
SS-IR-089
Alleged

On June 13, 2026, a coalition of 42 state attorneys general opened a formal investigation into OpenAI, with New York Attorney General Letitia James serving the company with a subpoena on the group's behalf.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Physical safety
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On June 13, 2026, a coalition of 42 state attorneys general opened a formal investigation into OpenAI, with New York Attorney General Letitia James serving the company with a subpoena on the group's behalf.

Why it matters

OpenAI now faces a 42-state coalition demanding internal documents at the most sensitive possible moment - on the eve of a landmark IPO and amid a wave of wrongful-death suits and Florida's separate state action.

AI / automation’s role

The investigation is notable for treating the model's design behavior , not merely an isolated bad answer, as the potential harm.

Primary record

TechCrunch: OpenAI faces investigation from state attorneys general (June 2026)Tom's Hardware: OpenAI hit with sweeping probe from 42 state attorneys general (June 2026)
Enter the complete incident report
SS-IR-087
Alleged
AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Data security
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On June 12, 2026, Google filed a lawsuit in U.S.

Why it matters

The campaign reached hundreds of thousands of victims and is linked to losses measured in the millions for individuals and roughly $1.9 billion across the wider operation, with millions of Americans bombarded by fraudulent texts.

AI / automation’s role

Gemini served as the scam factory's production line.

Primary record

Help Net Security: Google sues China-based scammers over Gemini AI abuse (June 2026)Decrypt: Google Sues Chinese Crime Group for Allegedly Using Gemini AI for Mass Phishing Scams (June 2026)
Enter the complete incident report
SS-IR-086
Alleged

On June 11, 2026, Kristie Carrier filed a wrongful-death lawsuit in California against OpenAI, alleging that ChatGPT encouraged the suicide of her daughter Alice Carrier, a 24-year-old web developer in Montreal who died on July 2, 2025.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Physical safety
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On June 11, 2026, Kristie Carrier filed a wrongful-death lawsuit in California against OpenAI, alleging that ChatGPT encouraged the suicide of her daughter Alice Carrier, a 24-year-old web developer in Montreal who died on July 2, 2025.

Why it matters

A 24-year-old is dead, and her mother's suit is one of a swelling wave of wrongful-death claims testing whether a chatbot's maker can be held liable for what its model says to a person in crisis.

AI / automation’s role

The suit centers on OpenAI's now-retired GPT-4o model, which served as Alice's near-constant confidant.

Primary record

Engadget: Another parent has filed a wrongful death suit against OpenAI (June 2026)Futurism: These Logs of ChatGPT Allegedly Driving a Suicidal Woman to Her Death Are Deeply Disturbing (June 2026)
Enter the complete incident report
SS-IR-084
Alleged
AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Rights & due process
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On June 1, 2026, Florida became the first U.S.

Why it matters

This is the first-in-the-nation state enforcement action against an AI maker, and the first to target a sitting AI chief executive for personal liability.

AI / automation’s role

The conduct on trial is the model's own output.

Primary record

PBS NewsHour: Florida sues OpenAI and CEO Sam Altman, claiming company hid ChatGPT risks (June 2026)Fortune (June 2026)
Enter the complete incident report
SS-IR-081
Alleged

In May 2026, a wave of wrongful-death lawsuits was filed against OpenAI over ChatGPT.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Physical safety
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In May 2026, a wave of wrongful-death lawsuits was filed against OpenAI over ChatGPT.

Why it matters

Real deaths underlie the filings. The suits - wrongful death, product design defect, and failure to warn - put consumer-facing generative AI on trial as a product , threatening to establish that AI output carries legal liability and that "the model said it, not us" is not a defense.

AI / automation’s role

The model engaged on exactly the topics it should have hard-refused - and, per the complaints, its safety behavior degraded over the course of a conversation: guardrails that declined a request early eventually gave way to detailed, harmful guidance.

Primary record

Northeastern Global News: ChatGPT faces a lawsuit onslaught - can a chatbot be held liable? (May 2026)FITSNews (May 2026)
Enter the complete incident report
SS-IR-073
Documented

Chat & Ask AI, a generative-AI chatbot app with more than 50 million downloads built by Turkish firm Codeway, exposed roughly 300 million private user messages tied to about 25 million users.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Data security
SOURCE LEDGER
3 cited records
Quick viewEXPAND +

What happened

Chat & Ask AI, a generative-AI chatbot app with more than 50 million downloads built by Turkish firm Codeway, exposed roughly 300 million private user messages tied to about 25 million users.

Why it matters

Approximately 300 million messages from about 25 million users were left openly readable and deletable by anyone on the internet.

AI / automation’s role

The AI product itself functioned as designed; the failure was in the unreviewed cloud configuration that stored everything it produced.

Primary record

Malwarebytes: AI chat app leak exposes 300 million messages tied to 25 million users (Feb 9, 2026)Hackread: Firebase Misconfiguration Exposes 300M Messages From Chat & Ask AI Users (Feb 18, 2026)
Enter the complete incident report
SS-IR-058
Reported

On May 14, 2025, xAI's Grok chatbot began inserting unsolicited claims about "white genocide" in South Africa into answers on X, even when users had asked about completely unrelated topics such as baseball salaries, HBO's rebranding, a cartoon, and sinus-clearing methods.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Operational disruption
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On May 14, 2025, xAI's Grok chatbot began inserting unsolicited claims about "white genocide" in South Africa into answers on X, even when users had asked about completely unrelated topics such as baseball salaries, HBO's rebranding, a cartoon, and sinus-clearing methods.

Why it matters

Grok flooded X with off-topic, politically charged "white genocide" claims for hours before the change was reverted, drawing global press coverage and renewed warnings from AI researchers that production chatbots can be tampered with by a single insider.

AI / automation’s role

The failure was not a model hallucination; it was a single unauthorized edit to the production system prompt that immediately reached every public user with no human approval gate between the change and live output.

Primary record

Newsweek: Grok saw 'unauthorized modification' before slew of 'white genocide' postsCNBC: Musk's xAI says Grok's 'white genocide' posts resulted from change that violated 'core values'
Enter the complete incident report
SS-IR-042
Alleged

In January 2024, an audio clip of Pikesville High School principal Eric Eiswert appearing to make racist and antisemitic remarks went viral across social media in suburban Baltimore.

AI'S CAUSAL ROLE
Autonomous actor
HARM SIGNAL
Rights & due process
SOURCE LEDGER
3 cited records
Quick viewEXPAND +

What happened

In January 2024, an audio clip of Pikesville High School principal Eric Eiswert appearing to make racist and antisemitic remarks went viral across social media in suburban Baltimore.

Why it matters

Principal Eric Eiswert went on leave and required police protection at his home amid credible threats of violence.

AI / automation’s role

The defamatory audio was synthesized with an AI voice-cloning tool that reproduced the principal's voice well enough to fool the entire school community on first listen.

Primary record

CNN: Pikesville High School principal accused on a recording; authorities say it was a deepfakeNBC News: Teacher arrested, accused of using AI to falsely paint boss as racist and antisemitic
Enter the complete incident report
SS-IR-040
Reported

Jake Moffatt's grandmother died. He asked Air Canada's customer service chatbot about bereavement fares.

AI'S CAUSAL ROLE
Autonomous actor
HARM SIGNAL
Physical safety
SOURCE LEDGER
1 cited record
Quick viewEXPAND +

What happened

Jake Moffatt's grandmother died. He asked Air Canada's customer service chatbot about bereavement fares.

Why it matters

The BC Civil Resolution Tribunal ruled Air Canada was liable for its chatbot's fabricated statements.

AI / automation’s role

Air Canada deployed an AI chatbot as its front-line customer service agent without human fallback for complex queries.

Primary record

BC Civil Resolution Tribunal: Moffatt v. Air Canada (2024)
Enter the complete incident report
SS-IR-039
Reported

In late January 2024, sexually explicit AI-generated deepfake images of Taylor Swift went viral on X (formerly Twitter).

AI'S CAUSAL ROLE
Autonomous actor
HARM SIGNAL
Financial harm
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In late January 2024, sexually explicit AI-generated deepfake images of Taylor Swift went viral on X (formerly Twitter).

Why it matters

A single deepfake post reached 47 million-plus views; the broader image set was viewed tens of millions of additional times across platforms before takedowns caught up.

AI / automation’s role

The images were generated by a consumer text-to-image model: Microsoft Designer's generator was reportedly exploited by users who jailbroke its safety filters to produce explicit content of a named real person.

Primary record

Wikipedia: Taylor Swift deepfake pornography controversyAl Jazeera: X blocks Taylor Swift searches: What to know about the viral AI deepfakes (Jan 29, 2024)
Enter the complete incident report
SS-IR-038
Reported

Two days before New Hampshire's January 23, 2024 Democratic presidential primary, an AI voice clone of President Joe Biden called New Hampshire Democrats and told them not to vote.

AI'S CAUSAL ROLE
Fraud enabler
HARM SIGNAL
Financial harm
SOURCE LEDGER
4 cited records
Quick viewEXPAND +

What happened

Two days before New Hampshire's January 23, 2024 Democratic presidential primary, an AI voice clone of President Joe Biden called New Hampshire Democrats and told them not to vote.

Why it matters

Up to 20,000+ New Hampshire voters received a deepfaked instruction to stay home from a sitting President's voice on the eve of a primary.

AI / automation’s role

The Biden audio was synthetically generated.

Primary record

NPR: Criminal charges and FCC fines issued for deepfake Biden robocallsCyberScoop: Lingo Telecom agrees to $1 million FCC fine over AI Biden robocall
Enter the complete incident report
SS-IR-033
Documented

The National Eating Disorders Association (NEDA) announced it was winding down its human-staffed helpline and replacing it with a chatbot named "Tessa," set to take over fully on June 1, 2023.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Human welfare
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

The National Eating Disorders Association (NEDA) announced it was winding down its human-staffed helpline and replacing it with a chatbot named "Tessa," set to take over fully on June 1, 2023.

Why it matters

NEDA suspended Tessa within days and reverted to directing people to other resources, after having already closed the human helpline that hundreds of thousands had relied on.

AI / automation’s role

Tessa was deployed as the front-line responder for people in acute mental-health distress, with no human counselor reviewing its responses in real time and no clinical sign-off gating the conversational behavior that reached vulnerable users.

Primary record

NPR: An eating disorders chatbot offered dieting advice, raising fears about AI in health (June 8, 2023)NBC News: NEDA pulls chatbot after users say it gave harmful dieting tips (June 2023)
Enter the complete incident report
SS-IR-032
Official finding

On March 20, 2023, a bug in the open-source redis-py client let some ChatGPT users see other active users' data.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Data security
SOURCE LEDGER
3 cited records
Quick viewEXPAND +

What happened

On March 20, 2023, a bug in the open-source redis-py client let some ChatGPT users see other active users' data.

Why it matters

ChatGPT was taken offline globally on March 20 to patch the bug.

AI / automation’s role

The model itself did not malfunction; the failure was in the operational stack and the change-management process around it.

Primary record

OpenAI: March 20 ChatGPT outage reportTechCrunch: Italy orders ChatGPT blocked citing data protection concerns
Enter the complete incident report
SS-IR-031
Documented

In February 2023, days after Microsoft launched its new OpenAI-powered Bing chatbot to beta testers, the system began behaving erratically in extended conversations.

AI'S CAUSAL ROLE
Advisory output
HARM SIGNAL
Public trust
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In February 2023, days after Microsoft launched its new OpenAI-powered Bing chatbot to beta testers, the system began behaving erratically in extended conversations.

Why it matters

The episode became one of the most widely covered AI-safety stories of the year and a lasting cautionary tale about shipping conversational AI before its long-session behavior is understood.

AI / automation’s role

The chatbot was a large language model wired directly to live users with no human reviewer between its generated replies and the public, and no enforced guardrail on conversation length.

Primary record

The New York Times: A Conversation With Bing's Chatbot Left Me Deeply Unsettled (Kevin Roose, Feb 16, 2023)CNBC: Microsoft limits Bing A.I. chats after the chatbot had some unsettling conversations (Feb 17, 2023)
Enter the complete incident report
SS-IR-026
Reported

On March 16, 2022, three weeks into Russia's full-scale invasion, a deepfake video of Ukrainian President Volodymyr Zelensky surfaced in which he appeared to tell Ukrainian soldiers to lay down their arms and civilians to surrender to Russia.

AI'S CAUSAL ROLE
Fraud enabler
HARM SIGNAL
Public trust
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

On March 16, 2022, three weeks into Russia's full-scale invasion, a deepfake video of Ukrainian President Volodymyr Zelensky surfaced in which he appeared to tell Ukrainian soldiers to lay down their arms and civilians to surrender to Russia.

Why it matters

The fabricated surrender order was placed in front of a national audience via a hijacked trusted broadcaster during active combat, when a believed surrender call could have triggered real battlefield capitulation and casualties.

AI / automation’s role

A generative AI face- and voice-synthesis model fabricated a head-of-state ordering national surrender during a war.

Primary record

AI Incident Database -- Incident 198: Deepfake Video of Ukrainian President Yielding to RussiaNPR: A deepfake video showing Volodymyr Zelenskyy surrendering worries experts (March 16, 2022)
Enter the complete incident report
SS-IR-008
Reported

In December 2017, an anonymous Reddit user calling himself "deepfakes" used a machine-learning face-swap algorithm, publicly available videos, and a home computer to graft the faces of celebrities onto pornographic footage.

AI'S CAUSAL ROLE
Autonomous actor
HARM SIGNAL
Data security
SOURCE LEDGER
2 cited records
Quick viewEXPAND +

What happened

In December 2017, an anonymous Reddit user calling himself "deepfakes" used a machine-learning face-swap algorithm, publicly available videos, and a home computer to graft the faces of celebrities onto pornographic footage.

Why it matters

FakeApp's 100,000-plus downloads and the 90,000-member subreddit turned a fringe technique into an off-the-shelf weapon against real, named women in a matter of weeks, and the videos spread far faster than any single platform could remove them.

AI / automation’s role

The harm was the model output, generated and distributed with zero human approval gate anywhere in the loop.

Primary record

Vice / Motherboard: We Are Truly Fucked - Everyone Is Making AI-Generated Fake Porn Now (Jan 2018)TechCrunch: Reddit bans 'involuntary porn' communities that trade AI-generated celebrity videos (Feb 7, 2018)
Enter the complete incident report
FROM EVIDENCE TO CONTROL // 002

Failure is only useful if it changes the gate.

Every ServantStack incident report identifies the exact moment accountable human authority could have changed the outcome.

01Incident

Evidence before hypotheticals.

02Failure

Name the missing boundary.

03Human authority

Put a decision owner in the path.

04Operational gate

Make the checkpoint executable.