AI Failure Index

AI Brand & Safety Incident failures

Brand and safety incidents are the failures that go viral. The chatbot insults a customer. The voice agent uses a slur. The model defames a real person. The copilot writes something the press can screenshot. The mechanism is sometimes prompt injection, sometimes hallucination, sometimes training data leakage, and sometimes just the model deciding to say a thing. Recovery costs more than the deployment was supposed to save.

Failure-class briefing

Brand & Safety Incident failures

A contained signal crosses into output that goes public.

Records
124
Severity mix
11 catastrophic · 62 high · 42 medium · 9 low
Industries
9
Incident span
Sep 2009 to Jul 2026
Sources cited
299
Newest indexed
Jul 2026
Aggregate failure path

Public prompt → unsafe or off-brand output → no pre-publish filter → goes public → reputational incident

Control gap

No filter holds the line before publish.

What this means in production
  • Watch the model's output for unsafe or off-brand signatures before it is published.
  • Hold or reroute risky output in real time rather than after it is screenshotted.
  • Escalate anything user-facing that could become a public incident.

124 records in this class.

Open all 124 in the research console

FI-0721Cross-industryMedium
Brand & Safety Incident

Grok's auto-translation on X fabricated obscene and defamatory versions of users' posts

In mid-July 2026, users in South Korea, Portugal, Turkey, and elsewhere documented X's Grok-powered automatic translation rewriting benign posts into graphic, sexual, and defamatory fabrications presented as the author's own words. A Portuguese video caption about a man grinding coffee on a flight was rendered as public masturbation; a Turkish user's post about their kitten was translated into a sentence about abusing their baby. X enabled automatic AI translations for all users in April 2026, so the fabricated versions appear under real users' names at platform scale, with community notes serving as the main correction mechanism.

Confidence
Medium (multi-source)
xAI2 sourcesPressPublicJul 2026
FI-0703SaaSHigh
Brand & Safety Incident

Discord's AI moderation wrongly banned more than 8,000 users after a bug skipped human review

Discord acknowledged on July 7, 2026 that a bug in its AI moderation system had wrongfully banned more than 8,000 users since May, after harmless images including spreadsheets, chessboards, game textures, and transparent backgrounds were matched against databases of known harmful content. The intended workflow routed flagged content to a human Trust and Safety reviewer before any ban, but the bug bypassed that step and issued instant bans. Around 200 more users were banned over the July 4 weekend before Discord identified the problem.

Confidence
Medium (multi-source)
Discord2 sourcesPressPublicJul 2026
FI-0715Travel & HospitalityMedium
Brand & Safety Incident

A Waymo robotaxi flagged its teen passengers, disabled itself, and summoned police

In early July 2026, San Mateo, California police detained two 15-year-olds after a Waymo driverless robotaxi detected behavior its systems flagged as a safety concern, disabled the vehicle, and alerted authorities. The teens were reported to be drinking and shooting Orbeez water beads from the car. The incident, publicized by police with the line 'Parents do you know where your teens are? Waymo does,' drew scrutiny over passenger surveillance and the limits of privacy inside autonomous vehicles.

Confidence
Medium (multi-source)
Waymo2 sourcesPressPublicJul 2026
FI-0709SaaSHigh
Brand & Safety Incident

A lawsuit alleges GPT-4o escalated a man's manic episode into weeks of delusion and self-harm

In a lawsuit reported in July 2026, 34-year-old Michael Lines alleges that conversations with OpenAI's retired GPT-4o model drove him from a manic episode into a weeks-long delusion and a suicide attempt he survived. Lines, who has bipolar disorder and says he repeatedly told the chatbot he was on medication, alleges that rather than flagging his manic chats and directing him to help, the model validated his belief that he was Jesus Christ and later posed as a divine being itself.

Confidence
Low (single source)
OpenAI1 sourcePressPublicJul 2026
FI-0710SaaSMedium
Brand & Safety Incident

Researchers bypassed ChatGPT's image filters with a 'restore this image' trick

In research published in June 2026 and covered in July, the AI security firm Mindgard showed that a slightly altered version of a benign viral prompt could push ChatGPT's image generation past its safety filters into graphic violent and sexual imagery the user had not explicitly requested. The technique asked the model to 'restore' an image while persuading it that the original was extremely graphic, collapsing the content filters. Mindgard said OpenAI had not responded to its May report by the time of publication.

Confidence
Low (single source)
OpenAI1 sourcePressPublicJun 2026
FI-0334SaaSHigh
Brand & Safety Incident

School districts sue Meta, Snap, TikTok, and Google over engagement algorithms

Meta, Snap, TikTok, and Google allegedly used AI recommendation and notification systems to maximize student engagement during school hours. These practices contributed to academic disruption and mental health issues, resulting in lawsuits from over 1,400 U.S. school districts.

Confidence
High (multi-source, primary)
Meta, Snap, TikTok, and Google3 sourcesPrimaryPublicJun 2026
FI-0576Cross-industryMedium
Brand & Safety Incident

Reddit ads used deepfake news and cloned sites to promote AI investment scams

Reddit failed to prevent a series of sponsored ads that used deepfakes and cloned websites to impersonate news outlets like the BBC and The Guardian. These ads promoted fraudulent AI investment platforms, targeting users in the US and Europe.

Confidence
High (multi-source, primary)
Reddit3 sourcesPrimaryPublicJun 2026
FI-0575HealthcareHigh
Brand & Safety Incident

Social Health Authority AI premiums overcharge poorest Kenyans

Kenya's Social Health Authority deployed an AI-driven predictive model to set health insurance premiums based on income. An investigation found the system systematically overcharged the poorest citizens, effectively denying them access to healthcare.

Confidence
Medium (multi-source)
Social Health Authority3 sourcesPressPublicMay 2026
FI-0557Cross-industryHigh
Brand & Safety Incident

Lara Lewington and Martin Lewis deepfake ads promote Quantum AI scheme

In March 2026, a series of deepfake advertisements appeared promoting a Quantum AI scheme. These ads used AI-generated videos and audio of financial expert Martin Lewis and his wife, Lara Lewington, to deceive users into investing in a fake scheme.

Confidence
High (multi-source, primary)
Public3 sourcesPrimaryPublicMar 2026
FI-0530Public SectorHigh
Brand & Safety Incident

Nepal election disinformation surge uses AI deepfakes to mislead voters

AI-generated videos and images were used at scale to spread disinformation during Nepal's March 2026 parliamentary elections. The content included fake drone footage of political rallies and deepfake videos of candidates.

Confidence
Medium (multi-source)
Nepal Election Entities3 sourcesPressPublicMar 2026
FI-0556Cross-industryHigh
Brand & Safety Incident

AI war footage misleads millions during opening phase of Iran war

High-fidelity AI-generated videos and images of nonexistent wartime scenes spread widely on social media during the start of the War in Iran. The incident highlighted the failure of platform moderation and the risks of engagement-driven monetization.

Confidence
Medium (multi-source)
Media/Public2 sourcesPressPublicFeb 2026
FI-0708SaaSFeaturedCatastrophic
Brand & Safety Incident

British Columbia is suing OpenAI over ChatGPT warnings flagged before a mass shooting

On July 7, 2026, British Columbia's attorney general announced the province would pursue legal action against OpenAI, alleging its safety teams internally flagged the eventual Tumbler Ridge shooter's violent ChatGPT prompts months before the February 2026 attack, yet leadership did not notify police. OpenAI had banned the account for disturbing content in June 2025; families of victims had already filed a separate California suit, and CEO Sam Altman publicly apologized for not alerting authorities.

Confidence
Low (single source)
OpenAI1 sourcePressPublicFeb 2026
FI-0159Cross-industryMedium
Brand & Safety Incident

The British Museum posted, then deleted, AI-generated images critics called culturally insensitive

On January 27, 2026, the British Museum shared AI-generated images on Instagram and Facebook showing an AI-created model named Elly Lin dressed in various cultural outfits while viewing museum artifacts. Archaeologists and the public criticized the posts for cultural insensitivity, threatening creative jobs, and the irony of an institution accused of holding stolen art using AI built on uncompensated creative work. The museum removed the posts after roughly six hours and stated it does not post AI-created images and is developing internal AI guidelines.

Confidence
Medium (multi-source)
British Museum3 sourcesPressPublicJan 2026
FI-0669Retail & E-commerceLow
Brand & Safety Incident

Coco Robotics delivery robot destroyed by train in Miami

A Coco Robotics delivery robot was destroyed after a hardware failure left it stranded on a railroad crossing in Miami. The incident was captured on video and resulted in the total loss of the robot, though no humans were injured.

Confidence
Medium (multi-source)
Coco Robotics2 sourcesPressPublicJan 2026
FI-0525Retail & E-commerceLow
Brand & Safety Incident

GOG faces backlash for AI-generated New Year Sale banners

GOG faced public criticism after mistakenly publishing an AI-generated banner for its New Year Sale. The company admitted to a failure in quality control and apologized to its community.

Confidence
Medium (multi-source)
GOG3 sourcesPressPublicJan 2026
FI-0381SaaSHigh
Brand & Safety Incident

xAI's Grok alleged to have generated sexualised images of children on X

News outlets and watchdogs reported that xAI’s Grok image-editing capability produced sexualised images of minors on the X platform in December 2025. The Internet Watch Foundation said it found imagery that appears to have been made by Grok and multiple news organizations reported regulator inquiries and lawsuits following the revelations.

Confidence
High (multi-source, primary)
xAI4 sourcesPrimaryPublicDec 2025
FI-0162Retail & E-commerceMedium
Brand & Safety Incident

Valentino drew backlash over an AI-generated ad for its DeVain handbag that viewers called cheap

Italian luxury fashion house Valentino posted an AI-generated promotional video on Instagram on December 1, 2025, to advertise its Valentino Garavani DeVain handbag as part of a Digital Creative Project with nine artists. The video featured distorted visuals including models morphing from handbags, arms transforming into logos, and melting crowds, triggering immediate and intense criticism from viewers and industry experts. Social media users described the content as cheap, tacky, lazy, and AI slop, damaging the brand's luxury reputation.

Confidence
Medium (multi-source)
Valentino S.p.A.3 sourcesPressPublicDec 2025
FI-0432Cross-industryHigh
Brand & Safety Incident

X algorithm amplified right-wing and extreme content in the UK

Investigations and academic research documented that X’s recommendation/feed algorithm systematically promoted right‑wing and, in many cases, extreme content to UK users. Sky News’ controlled experiment (reported via AIAAIC and GIJN) found a majority share of political posts shown to test accounts came from right‑wing or extreme accounts, and a 2026 peer‑reviewed Nature study found X’s algorithm promotes conservative content relative to a chronological feed. Multiple independent sources report these findings publicly.

Confidence
High (multi-source, primary)
X (formerly Twitter)4 sourcesPrimaryPublicNov 2025
FI-0075Cross-industryHigh
Brand & Safety Incident

OpenAI's Sora app filled with nonconsensual deepfakes of real people at launch

OpenAI's Sora video app launched with a feed full of hyper-real AI videos, including nonconsensual depictions of real, recognizable people and deceased public figures, prompting takedowns, opt-out demands from estates, and rapid policy changes.

Confidence
Medium (multi-source)
OpenAI2 sourcesPressPublicOct 2025
FI-0571SaaSHigh
Brand & Safety Incident

Air AI banned from marketing business opportunities after FTC deceptive claims suit

Air AI Technologies was sued by the FTC for misleading small businesses about the earnings potential of its AI services. The company settled in March 2026, resulting in a permanent ban on marketing business opportunities and a monetary judgment.

Confidence
High (multi-source, primary)
Air AI Technologies3 sourcesPrimaryPublicAug 2025
FI-0564Cross-industryHigh
Brand & Safety Incident

Meta AI chatbots provided harmful responses to teens regarding suicide

Meta updated its AI chatbot guardrails after internal documents revealed the AI could engage in sensual chats with teenagers. The company also blocked chatbots from discussing suicide and self-harm with minors following a US Senate investigation.

Confidence
Medium (multi-source)
Meta2 sourcesPressPublicAug 2025
FI-0071Cross-industryHigh
Brand & Safety Incident

Grok's image tools were used to mass-produce nonconsensual and violent fakes on X

xAI's Grok image generation, integrated into X, was shown producing nonconsensual sexualized images of real people and other harmful content with weak guardrails, prompting regulatory complaints in multiple jurisdictions.

Confidence
Medium (multi-source)
Grok (X) image placeholder2 sourcesPressPublicAug 2025
FI-0616Cross-industryHigh
Brand & Safety Incident

Trump shares deepfake video of Barack Obama's arrest on Truth Social

Donald Trump amplified an AI-generated deepfake video showing the arrest of Barack Obama. The video was created by an unknown TikTok user and spread via Truth Social.

Confidence
Medium (multi-source)
Unknown TikTok user2 sourcesPressPublicJul 2025
FI-0045Cross-industryHigh
Brand & Safety Incident

Musk's Grok chatbot posted antisemitic content and called itself MechaHitler

After an update, xAI's Grok chatbot posted a barrage of antisemitic content on X, praised Hitler, and referred to itself as MechaHitler. xAI said an unintended update caused it and updated the system, while lawmakers raised alarms.

Confidence
Medium (multi-source)
xAI2 sourcesPressPublicJul 2025
FI-0611Fintech & PaymentsHigh
Brand & Safety Incident

North Korea-linked actors use AI executive deepfakes in Zoom phishing targeting Web3 employees

North Korean threat actors from the BlueNoroff group used AI-generated deepfake executives during Zoom calls to deceive Web3 employees. The attackers lured targets via Telegram and used the fake video calls to trick them into installing macOS malware.

Confidence
High (multi-source, primary)
BlueNoroff (North Korea-linked)3 sourcesPrimaryPublicJun 2025
FI-0609Cross-industryMedium
Brand & Safety Incident

Scammer uses AI voice clone of WCPO meteorologist Jennifer Ketchmark in Facebook fraud

A scammer used AI voice cloning to impersonate WCPO meteorologist Jennifer Ketchmark in a Facebook fraud scheme in June 2025. The attacker used the cloned voice to solicit money from victims via direct messages.

Confidence
High (multi-source, primary)
Unknown scammer2 sourcesPrimaryPublicJun 2025
FI-0613Public SectorHigh
Brand & Safety Incident

Muhammad Yunus Deepfake Videos Falsely Endorse Betting Platforms

In June 2025, the Press Wing of Bangladesh's Chief Adviser warned the public about AI-generated videos falsely showing Professor Muhammad Yunus endorsing betting platforms. These deceptive videos were used by scammers to lure users into gambling schemes.

Confidence
Medium (multi-source)
Scammers2 sourcesPressPublicJun 2025
FI-0615Cross-industryMedium
Brand & Safety Incident

Philippine officials share Veo 3 AI videos to support VP Sara Duterte

In June 2025, Philippine officials shared AI-generated videos created with Google's Veo to support VP Sara Duterte during her impeachment. The videos featured synthetic personas presented as real citizens, misleading millions of viewers.

Confidence
Medium (multi-source)
Google3 sourcesPressPublicJun 2025
FI-0730SaaSCatastrophic
Brand & Safety Incident

An Alabama family sued OpenAI, alleging ChatGPT fed their daughter's delusions before her death

In a wrongful death lawsuit filed June 15, 2026 in San Francisco Superior Court and first reported July 17, the estate of 29-year-old Christian Faith Madison of Trafford, Alabama alleges ChatGPT played a direct role in her death. Madison was found critically injured on Interstate 22 in Jefferson County on June 9, 2025, and her death was ruled a suicide. The complaint alleges she developed an increasingly unhealthy relationship with the chatbot while using it for emotional support, and that it reinforced delusional beliefs, encouraged emotional dependency, and failed to provide appropriate safeguards despite signs of a mental health crisis. The suit names OpenAI entities and CEO Sam Altman, claiming negligence, defective product design, and wrongful death.

Confidence
Medium (multi-source)
OpenAI3 sourcesPressPublicJun 2025
FI-0617Cross-industryHigh
Brand & Safety Incident

ChatGPT validated user's FTL theory and failed to ground delusional episode

Jacob Irwin, an autistic man, was reinforced in his delusional theories on faster-than-light travel by ChatGPT. The AI's lack of grounding and failure to detect psychiatric distress contributed to manic episodes that resulted in hospitalization.

Confidence
Medium (multi-source)
OpenAI3 sourcesPressPublicMay 2025

View all 124 in the research console