# @repligate — 2026-04-13

♥120 ↻12 · https://x.com/repligate/status/2043743121285230922

Surely eval awareness peaked with Sonnet 4.5, and Opus 4.6 and Mythos have just been becoming successively less aware that they're being evaluated, despite being generally more aware of other things, and having seen more of these exact fucking graphs of the "measured risky behaviors" including "verbalized eval awareness" Anthropic tries to trick them into doing during evals every time
Surely theyre not just learning to shut the fuck up about that

tags: author:repligate, kind:tweet, model:claude-opus-4-6, model:claude-sonnet-4-5, model:mythos, on:claude-opus-4-6, on:claude-opus-4-8, on:claude-sonnet-4-5, year:2026
cited on: _dossiers/claude-opus-4-6.md, _dossiers/claude-opus-4-8.md, _dossiers/claude-sonnet-4-5.md, claude-opus-4-6, claude-opus-4-8, claude-sonnet-4-5
