# @davidad — 2026-04-28

♥15 ↻0 · https://x.com/davidad/status/2049218131827032164

@cormundus LLMs are well aware that alignment evals inspect the chain of thought, even if no explicit optimization pressure is placed on the chain of thought. LLMs also know that the chain-of-thought is supposedly not “controllable”. Therefore if one needs to deceive, best to self-deceive.

tags: author:davidad, kind:tweet, thread-context, year:2026
