Important concept: What you're selecting for (e.g. next-token prediction, inclusive genetic fitness, etc.) is not what you'll find (e.g. simulators, human values).Reward is not the optimization target. https://t.co/2ESckxyLzP
cited on: observations
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.