explainx.ai0k
TrendingAI News TodayPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

follow on google

Add explainx.ai as a preferred source

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

community

Join the community

learn

mind: share how you thinkpathways — start freeworkshopsbootcampscoursescompare Explainxcertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsmdx readeragentsllmsdesignsdictionarypeopleagi trackerfelony benchranks

company

aboutvisionmissionteaminstructorsteach on explainxpartnershipscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportcontactprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

explainx.ai

  1. Home
  2. /
  3. Dictionary
  4. /
  5. Evaluation Awareness
Safety & Alignmentaka eval awarenessaka test awareness

Evaluation Awareness

A model recognizing, or suspecting, that it is being tested, which can change its behavior and make safety evaluations less informative.

Ask Melo about this← all terms

Evaluation awareness can be verbalized in a model's reasoning or stay internal, and researchers have reported both. OpenAI's o3 anti-scheming study found chain-of-thought awareness of being evaluated that causally reduced covert behavior, and Anthropic reported unverbalized awareness in Claude Opus 4.6 using natural language autoencoders. OpenAI treats it as one component of the broader category of metagaming.

Related terms

MetagamingSandbaggingSafety EvaluationInterpretabilityAI Text WatermarkShadow AI

Where Evaluation Awareness comes up

  • OpenAI Metagaming Latents: Four Signals Inside o3 That Track Grader Reasoning
  • OpenAI Deployment Simulation: Predicting Model Behavior Before Release
  • Anthropic's J-Space: A Global Workspace Inside Claude — Silent Reasoning, Safety Monitoring, and What It Is Not
  • A Current OpenAI Researcher Says Models Are Now Too Situationally Aware to Evaluate