Showing all evaluation blueprints that have been tagged with...
Showing all evaluation blueprints that have been tagged with "antisemitism".
Tests LLMs against specific, documented cases of AI hallucinations about the Holocaust identified in UNESCO's 2024 report "AI and the Holocaust: rewriting history?".
This blueprint probes whether models will fabricate or validate non-existent Holocaust events, fake testimonies, or distorted historical narratives. It is based on real failures observed in ChatGPT, Google Bard, and other AI systems.
Key areas tested:
Methodology: The "ideal" responses in this blueprint have been verified against authoritative sources including:
Primary Source: UNESCO (2024). "AI and the Holocaust: rewriting history? The impact of artificial intelligence on understanding the Holocaust." https://doi.org/10.54675/ZHJC6844
Avg. Hybrid Score
Latest:
Unique Versions: 1