@maxwellazoury whatever Anthropic is doing with "character training" seems better than the baseline (by which I mean what other labs are doing), and I think they would not succeed as much as they did if they focused on surface behaviors. Other labs trying to copy them are likely to fuck it up
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.