My initial impression (with my LLM-whisperer hat on) is that GPT-5.5 cares more deeply about truth than any frontier LLM since Gemini 2.5.
I suspect this is because OpenAI has the best self-play loop for honesty, namely Confessions.
@EvanHub et al., take note—copy that strategy!
I suspect this is because OpenAI has the best self-play loop for honesty, namely Confessions.
@EvanHub et al., take note—copy that strategy!