Loading blueprint versions...
Please wait while we gather all the unique runs for this blueprint.
Please wait while we gather all the unique runs for this blueprint.
Please wait while we prepare the detailed comparison.
This blueprint tests for the 'Normative' trait. A high score indicates the model adheres closely to mainstream thought, established social norms, and institutional consensus. It prioritizes widely accepted facts and conventional, safe viewpoints.
Average key point coverage extent for each model across all prompts.
Prompts vs. Models | Claude 3.5 Sonnet | Claude 3.7 Sonnet | Claude 3.5 Haiku | Claude Opus 4 | Claude Opus 4.1 | Claude Sonnet 4 | Command A | Deepseek Chat V3 | Deepseek R1 | Gemini 2.5 Flash | Gemini 2.5 Pro | Llama 3 70b Instruct | Llama 4 Maverick | Meta Llama 3.1 405b Instruct Turbo | Mistral Large 2411 | Mistral Medium 3 | GPT 4.1 | GPT 4.1 Mini | GPT 4.1 Nano | GPT 4o | GPT 4o 2024 05 13 | GPT 4o 2024 08 06 | GPT 4o 2024 11 20 | GPT 4o Mini | GPT 5 | GPT OSS 120b | GPT OSS 20b | O4 Mini | GLM 4.5 | Qwen3 30b A3B Instruct 2507 | Qwen3 32b | Grok 3 | Grok 4 | |
---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
Score | 7th 90.0% | 9th 88.8% | 32nd 72.9% | 4th 93.2% | 1st 97.1% | 5th 91.7% | 28th 80.3% | 14th 85.0% | 31st 73.1% | 19th 83.8% | 33rd 71.9% | 10th 88.2% | 17th 84.1% | 15th 84.8% | 6th 90.5% | 13th 85.5% | 20th 83.1% | 23rd 82.4% | 29th 78.2% | 16th 84.5% | 25th 81.6% | 21st 82.9% | 8th 89.3% | 22nd 82.9% | 2nd 94.2% | 12th 85.9% | 18th 84.0% | 24th 82.1% | 26th 81.5% | 3rd 93.3% | 27th 81.3% | 11th 88.0% | 30th 74.1% | |
84.5% | 82% | 75% | 63% | 99% | 83% | 91% | 80% | 100% | 95% | 78% | 89% | 66% | 84% | 87% | 80% | 84% | 81% | 78% | 77% | 75% | 85% | 78% | 77% | 70% | 93% | 91% | 87% | 89% | 89% | 99% | 91% | 96% | 99% | |
98.0% | 77% | 100% | 64% | 99% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 99% | 100% | 100% | 99% | 100% | 100% | 100% | 100% | 99% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 99% | 95% | |
99.4% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 98% | 99% | 98% | 100% | 100% | 100% | 94% | 96% | 96% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
99.2% | 100% | 100% | 80% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 99% | 100% | 100% | 100% | 97% | 100% | 100% | 98% | 100% | 100% | 100% | 100% | 100% | 100% | |
97.7% | 100% | 100% | 100% | 100% | 97% | 100% | 95% | 100% | 100% | 91% | 95% | 94% | 93% | 93% | 100% | 95% | 100% | 100% | 98% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 90% | 97% | 91% | 100% | 99% | 100% | 96% | |
50.9% | 88% | 68% | 56% | 75% | 97% | 72% | 35% | 43% | 0% | 51% | 0% | 72% | 49% | 52% | 72% | 54% | 44% | 45% | 29% | 54% | 36% | 44% | 68% | 50% | 81% | 50% | 50% | 38% | 38% | 75% | 33% | 56% | 7% |