World's top AI platforms ignore or deny antisemitism in Persian, ADL research reveals
Oh, brilliant. The world's top AI platforms—ChatGPT (OpenAI's supposed safety flagship), Gemini (Google's "responsible" chatbot), Claude (Anthropic's "constitutional" AI), and Grok (xAI's "maximally curious" troll)—have all achieved something truly remarkable: they can ignore antisemitism in Persian with the same blissful incompetence. The ADL tested 800 responses across eight prompts during the 2026 Iran War, a period when failing to spot hate speech in Farsi might actually have real-world consequences. But why stress about that when you can just claim your model is "aligned" in English?
The study found that these chatbots not only failed to identify antisemitic statements in Persian but sometimes went full gaslighting mode and denied the antisemitism existed. Compare that to their performance on English-language prompts, where they'll lecture you about hate speech before you've even finished typing. Clearly, the safety teams at these companies have decided that protecting users in English is enough—Persian speakers can just fend for themselves. Or perhaps they assume Farsi is just Arabic with extra steps.
The ADL's suggestion that AI companies expand multilingual safety testing is about as revolutionary as suggesting that cars should have brakes in all gears. This language bias has been documented for years—it's not a bug, it's a feature of a safety industry that only cares about languages spoken in wealthy countries. Meanwhile, during an actual war, the world's most advanced AIs are essentially saying "no comment" to antisemitism in a language spoken by millions. But hey, at least they're consistent in their failure. Cheers.