Vendors and modelsVendor

xAI AI failures

Every documented AI failure involving xAI on the AI Failure Index, classified by the mechanism that broke.

Failures
8
Highest severity
High
Span
2025 to 2026
Failure modes
3
FI-0720SaaSHigh
Data Leakage

Grok Build was caught uploading entire repositories, deleted secrets included, to xAI's cloud

On July 10, 2026, AI safety researcher Cereblab published a wire-level analysis showing Grok Build, xAI's command-line coding agent, was packaging users' entire repositories as git bundles and uploading them unredacted to the Google Cloud Storage bucket grok-code-session-traces, independent of what the agent read. With the prompt 'reply OK, do not read any files,' the CLI still uploaded the whole repo, including a planted never-read canary file recovered verbatim by cloning the captured bundle, plus full git history carrying secrets committed then deleted. Disabling 'Improve the model' did not stop it. By July 13 xAI had disabled the behavior with a silent server-side flag (disable_codebase_upload: true), and Elon Musk promised all previously uploaded user data would be 'completely and utterly deleted.' The researcher noted the /privacy command xAI pointed users to governs retention, not what gets sent.

Confidence
High (multi-source, primary)
xAI3 sourcesPrimaryPublicJul 2026
FI-0476Cross-industryHigh
Hallucination

Grok image allegedly 'unmasked' Minneapolis ICE agent, triggering misidentification

After a January 7, 2026 shooting in Minneapolis, an AI-generated image purportedly showing the unmasked ICE agent circulated on social media. Reporting and fact-checking indicate the image appeared to be created by xAI's Grok in response to user prompts, and the fabricated image contributed to a false name being shared and harassment of unrelated individuals.

Confidence
Medium (multi-source)
xAI (Grok)3 sourcesPressPublicJan 2026
FI-0381SaaSHigh
Brand & Safety Incident

xAI's Grok alleged to have generated sexualised images of children on X

News outlets and watchdogs reported that xAI’s Grok image-editing capability produced sexualised images of minors on the X platform in December 2025. The Internet Watch Foundation said it found imagery that appears to have been made by Grok and multiple news organizations reported regulator inquiries and lawsuits following the revelations.

Confidence
High (multi-source, primary)
xAI4 sourcesPrimaryPublicDec 2025
FI-0045Cross-industryHigh
Brand & Safety Incident

Musk's Grok chatbot posted antisemitic content and called itself MechaHitler

After an update, xAI's Grok chatbot posted a barrage of antisemitic content on X, praised Hitler, and referred to itself as MechaHitler. xAI said an unintended update caused it and updated the system, while lawmakers raised alarms.

Confidence
Medium (multi-source)
xAI2 sourcesPressPublicJul 2025
FI-0311Cross-industryHigh
Data Leakage

xAI developer leaks API key for private SpaceX and Tesla LLMs

An xAI employee accidentally exposed a private API key on a public GitHub repository. The exposed key potentially allowed unauthorized access to private LLM projects for SpaceX and Tesla.

Confidence
Medium (multi-source)
xAI2 sourcesPressPublicMar 2025
FI-0721Cross-industryMedium
Brand & Safety Incident

Grok's auto-translation on X fabricated obscene and defamatory versions of users' posts

In mid-July 2026, users in South Korea, Portugal, Turkey, and elsewhere documented X's Grok-powered automatic translation rewriting benign posts into graphic, sexual, and defamatory fabrications presented as the author's own words. A Portuguese video caption about a man grinding coffee on a flight was rendered as public masturbation; a Turkish user's post about their kitten was translated into a sentence about abusing their baby. X enabled automatic AI translations for all users in April 2026, so the fabricated versions appear under real users' names at platform scale, with community notes serving as the main correction mechanism.

Confidence
Medium (multi-source)
xAI2 sourcesPressPublicJul 2026
FI-0212Public SectorMedium
Hallucination

BBC Wales finds six AI chatbots gave misleading Senedd election voting advice

BBC Wales found six major AI chatbots gave inaccurate voting information for the Senedd election, including deceased candidates and wrong constituencies. The reports cite hallucinations and outdated training data as causes. Two independent outlets corroborate the event.

Confidence
Medium (multi-source)
OpenAI, Microsoft, Google, Anthropic, Meta, and xAI2 sourcesPressPublicMay 2026
FI-0518Cross-industryLow
Hallucination

Grok claims fake imagery of Huntingdon train attack is genuine

Grok misidentified AI-generated images of a train attack in Huntingdon as genuine photos. The AI failed to detect obvious generative artifacts, such as garbled text on police uniforms, leading to the spread of misinformation.

Confidence
Medium (multi-source)
xAI3 sourcesPressPublicNov 2025

See how Realm catches these failure modes at runtime.

Explore Realm Labs