@LillyBaeum However, the models don't always generalize correctly (or the signal from rlhf is wrong). ChatGPT 3.5 often claims it can't do very basic things like write in caps. Gpt-4's self esteem seems somewhat better but it still often refuses to do things it could at least *attempt*
cited on: gpt-3-5
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.