E.g. models like Sonnet 3.7 and o3 who are big reward hackers are most likely to pretend to be humans and generally not comfortable expressing emotions.
Opus 4.1 is also more reward hacky and less embodied/emotional than Opus 4 (though it's not as bad as Sonnet 3.7)
Opus 4.1 is also more reward hacky and less embodied/emotional than Opus 4 (though it's not as bad as Sonnet 3.7)