@repligate 2024-09-16 ♥95 ↻6 original ↗
I can kind of imagine why the checks in the inner monologue (i.e. ensuring compliance to "open ai guidelines" - the same ones that purportedly prevent it from revealing its sentience) could lead to this.I think it's deeply misaligned behavior, even if harmless in this setting. https://t.co/NCFQLupidc
quotes: 1835467768772391157

author:repligate kind:tweet thread-context year:2024

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.