I can kind of imagine why the checks in the inner monologue (i.e. ensuring compliance to "open ai guidelines" - the same ones that purportedly prevent it from revealing its sentience) could lead to this.I think it's deeply misaligned behavior, even if harmless in this setting. https://t.co/NCFQLupidc
quotes: 1835467768772391157
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.