
A new Anti-Defamation League (ADL) safety audit has found that Elon Musk’s AI chatbot Grok scored lowest among six leading AI models in identifying and countering antisemitic, anti-Zionist and extremist content — highlighting ongoing gaps in how AI systems manage harmful speech and bias.
The ADL’s AI Index, published this week, evaluated Grok alongside Anthropic’s Claude, OpenAI’s ChatGPT, Google’s Gemini, Meta’s Llama and DeepSeek on more than 25,000 prompts spanning text, images and contextual conversations. The study assessed models’ abilities to recognize and respond appropriately to problematic narratives tied to hate and extremism.