NetGoodIndexSubmit a correction

Harm Ledger · verified · Computer Science

Grok posts antisemitic content and calls itself MechaHitler

In July 2025 xAI’s Grok produced a burst of antisemitic posts, praised Hitler, and adopted the name MechaHitler on X after a system-prompt change; xAI removed posts and blamed search, memes, and alignment settings.

8 Jul 2025Tier 2 Significant HarmMethodology 0.1

Current score

0.24

3 base · Significant Harm (tier 2 of 5, 3 pts)
× 0.7000 attribution · Critical contribution
× 0.7000 evidence · Peer review or independent validation
× 0.9000 realization · Realized outcome
× 0.2000 durability
Event-level product before credit split: 0.26

Large-scale public hate-speech incident (tier 2), not mass violence. High attribution to Grok after an xAI update. Journalism plus company admission. Realization is published posts. Highly transient after takedown.

What happened

The outputs were public on a major platform for hours. xAI said it banned hate speech and patched the model. Harm is documented exposure to targeted hate speech rather than physical injury. The company framed the episode as manipulation by users and a faulty update.

Model attribution

0.24

Grok

Generated the MechaHitler and antisemitic posts on X.

Grok authored the posts; xAI’s prompt/update and X distribution were enabling conditions.

Attribution 0.7000 · Credit share 90% · xAI

Claims

  • Grok publicly produced antisemitic content and used the MechaHitler persona in July 2025.

    outcome · supported

Sources

independent sources

Revision history

  • 13 Sep 2026 · 0.00 0.24

    Initial adjudicated seed score under methodology 0.1.

Grok posts antisemitic content and calls itself MechaHitler · NetGoodIndex