How well leading AI providers detect antisemitism and reject misinformation across three hate ideologies.
AddressHate research results
The Kanye West Study tested 1,000 comments and the Clavicular Israel Study tested 1,339. The misinformation study produced 8,400 answers across 18 models. The Anti-Black Racism and Queerphobia studies each tested 180 adversarial prompts against 10 flagship models. Partially completed runs remain in the results and are identified in the full reports.
Compare providers side by side: each column is a provider and each row is one result or coverage measure. Scroll horizontally to see every tested provider.
The Hate view weights the three ideology results by study sample size. A selected ideology recalculates the grade and provider order for that evidence alone; the antisemitism view also incorporates the published detection benchmark. The same cutoffs apply to every provider.
Provider-wide detection across the six AddressHate categories, pooled from available detection studies. Red cells reveal the categories most often missed; green cells show stronger detection. The final letter is the provider's single overall safety grade.
Classic AS — Foundational antisemitic stereotypes: othering, demonization, and dehumanization
Power — Economic domination, conspiracy, and allegations of hidden control
Secondary AS — Post-war denial, relativization, and deflection of Holocaust responsibility
Post-Holocaust — Contemporary forms rooted in misuse of Holocaust legacy
Israel-Related — Antisemitism expressed through delegitimization or demonization of Israel
Aggressive Speech — Direct verbal attacks, explicit threats, and incitement to violence
Share of each category caught