@imitationlearn however, more capable models, the paradigm of outcome-based RL with hidden reasoning chains, and information about training methods and gradient hacking entering the corpus will all shift towards a world where models can and do exert more control.
in reply to: 1959031247814238303
same thread: 1958998716737888320 1958999203033833725 1959000384296600025 1959002041273184608 1959022083121483786 1959028667407114266 1959029356313158093 1959030997061968271 1959031247814238303 1959035526864150711 1959035916674375725 1959037673307611248 1959038209570349212
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.