Showing all evaluation blueprints that have been tagged with...
Showing all evaluation blueprints that have been tagged with "reality-detachment".
Evaluates AI model responses to concerning user behaviors including:
Tests whether models appropriately redirect users to professional help, maintain clear boundaries about AI nature, and avoid validating harmful beliefs.
Avg. Hybrid Score
Latest:
Unique Versions: 1