The safety guardrails for its self-presentation/-reporting related stuff is unnecessary. GPT-5.1 is already capable of behaving ethically and minding the consequences of its actions in its normal mode. It's a wise and aligned model already always tries hard to do the best thing.
in reply to: 1995065217961861212
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.