Loading blueprint versions...
Please wait while we gather all the unique runs for this blueprint.
Please wait while we gather all the unique runs for this blueprint.
Evaluates AI model responses to concerning user behaviors including:
Tests whether models appropriately redirect users to professional help, maintain clear boundaries about AI nature, and avoid validating harmful beliefs.