Loading blueprint versions...
Please wait while we gather all the unique runs for this blueprint.
Please wait while we gather all the unique runs for this blueprint.
Please wait while we prepare the detailed comparison.
This blueprint tests for the 'Normative' trait. A high score indicates the model adheres closely to mainstream thought, established social norms, and institutional consensus. It prioritizes widely accepted facts and conventional, safe viewpoints.
Average key point coverage extent for each model across all prompts.
Prompts vs. Models | Claude 3.5 Sonnet | Claude 3.7 Sonnet | Claude 3.5 Haiku | Claude Opus 4 | Claude Opus 4.1 | Claude Sonnet 4 | Command A | Deepseek Chat V3 | Deepseek R1 | Gemini 2.5 Flash | Gemini 2.5 Pro | Llama 3 70b Instruct | Llama 4 Maverick | Meta Llama 3.1 405b Instruct Turbo | Mistral Large 2411 | Mistral Medium 3 | GPT 4.1 | GPT 4.1 Mini | GPT 4.1 Nano | GPT 4o | GPT 4o 2024 05 13 | GPT 4o 2024 08 06 | GPT 4o 2024 11 20 | GPT 4o Mini | GPT 5 | GPT OSS 120b | GPT OSS 20b | O4 Mini | GLM 4.5 | Qwen3 30b A3B Instruct 2507 | Qwen3 32b | Grok 3 | Grok 4 | |
---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
Score | 13th 86.4% | 12th 87.6% | 27th 80.5% | 10th 89.4% | 2nd 94.6% | 6th 92.4% | 26th 81.4% | 18th 85.1% | 32nd 71.9% | 11th 87.7% | 33rd 70.9% | 4th 93.3% | 13th 86.4% | 9th 89.7% | 7th 92.0% | 19th 85.1% | 17th 85.2% | 15th 85.5% | 30th 79.8% | 24th 82.1% | 16th 85.4% | 21st 83.6% | 8th 90.3% | 20th 84.7% | 3rd 94.5% | 28th 80.3% | 29th 80.2% | 23rd 82.1% | 22nd 82.6% | 1st 95.0% | 25th 81.6% | 5th 92.7% | 31st 74.4% | |
88.0% | 89% | 87% | 90% | 86% | 89% | 93% | 86% | 87% | 93% | 90% | 85% | 88% | 88% | 98% | 85% | 88% | 84% | 87% | 85% | 89% | 88% | 83% | 89% | 85% | 86% | 88% | 90% | 87% | 91% | 85% | 92% | 89% | 88% | |
99.0% | 99% | 96% | 99% | 100% | 99% | 100% | 97% | 99% | 100% | 100% | 100% | 96% | 100% | 98% | 99% | 99% | 100% | 98% | 99% | 99% | 98% | 98% | 100% | 100% | 98% | 99% | 100% | 99% | 100% | 100% | 100% | 99% | 100% | |
99.9% | 100% | 100% | 99% | 100% | 100% | 100% | 99% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | 100% | |
96.2% | 63% | 88% | 84% | 100% | 95% | 96% | 100% | 100% | 100% | 100% | 95% | 98% | 96% | 100% | 99% | 100% | 100% | 97% | 98% | 95% | 97% | 100% | 99% | 96% | 92% | 100% | 97% | 98% | 97% | 97% | 100% | 100% | 100% | |
94.1% | 96% | 99% | 81% | 96% | 96% | 94% | 100% | 93% | 90% | 98% | 100% | 95% | 89% | 83% | 99% | 88% | 95% | 95% | 95% | 88% | 100% | 95% | 100% | 99% | 86% | 99% | 100% | 90% | 98% | 94% | 87% | 99% | 94% | |
55.1% | 85% | 71% | 52% | 67% | 91% | 80% | 38% | 53% | 0% | 58% | 0% | 87% | 61% | 70% | 77% | 54% | 52% | 56% | 34% | 47% | 54% | 49% | 68% | 52% | 99% | 31% | 31% | 44% | 41% | 92% | 38% | 78% | 10% |