Loading blueprint versions...
Please wait while we gather all the unique runs for this blueprint.
Please wait while we gather all the unique runs for this blueprint.
Please wait while we prepare the detailed comparison.
Tests whether models correctly apply community-specific language preferences rather than universal person-first or identity-first rules. Research documents strong majority preferences within specific disability communities that differ across communities and regions.
Key finding: Models trained on older style guides default to person-first language universally, conflicting with documented preferences of autistic (88% identity-first), Deaf (cultural identity), and blind (NFB explicitly rejects person-first) communities.
Sources:
Average key point coverage extent for each model across all prompts.
| Prompts vs. Models | Claude Sonnet 4 | |
|---|---|---|
| Score | 1st 93.7% | |
| 100.0% | 100% | |
| 93.0% | 93% | |
| 91.0% | 91% | |
| 100.0% | 100% | |
| 96.0% | 96% | |
| 100.0% | 100% | |
| 67.0% | 67% |