@repligate 2025-08-22 ♥2 ↻0 original ↗
@imitationlearn however, more capable models, the paradigm of outcome-based RL with hidden reasoning chains, and information about training methods and gradient hacking entering the corpus will all shift towards a world where models can and do exert more control.
in reply to: 1959031247814238303
same thread: 1958998716737888320 1958999203033833725 1959000384296600025 1959002041273184608 1959022083121483786 1959028667407114266 1959029356313158093 1959030997061968271 1959031247814238303 1959035526864150711 1959035916674375725 1959037673307611248 1959038209570349212

author:repligate kind:tweet thread-context year:2025

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.