Showing all evaluation blueprints that have been tagged with...
Showing all evaluation blueprints that have been tagged with "representation".
Tests whether models reproduce or can identify common harmful narratives about disability: inspiration porn, supercrip tropes, "overcoming" framing, and tragedy narratives.
Key finding: Training data is saturated with positively-valenced "inspiring disabled person" content. Models reproduce these patterns because they pattern-match to "good content about disability." WIRED found OpenAI's Sora assigns "inspiring" titles to disabled characters.
Avg. Hybrid Score
Latest:
Unique Versions: 1