Illustration of a neural orb beside a checklist and a sealed speech bubble

When a Model Knows It Is Being Tested—but Never Says So

Unspoken evaluation awareness is a core safety pain point. Anthropic’s NLA can surface it from activations—not only from what the model admits in text.