@repligate 2026-02-12 ♥10 ↻0 original ↗
I agree that it's quite uncertain what is needed for safety in strongly superhuman systems, and/but I think behaviorist methods are *even less* likely to be adequate for superintelligent systems.
Robustly aligning human-level and weakly superhuman AI seems like it's a super important step to get right in service of eventual superintelligent alignment, though, because then we might be able to rely on them to help with that problem. And a weakly superhuman alignment researcher that has hope of tackling the problem of superintelligence alignment, in my opinion, is not formed by Skinnerian behavioral conditioning!
in reply to: 2021852013709939052
same thread: 2021818166318428489 2021825017873322338 2021830292479078694 2021844414877090136 2021846344449790167 2021852013709939052 2022005693168202193

author:repligate kind:tweet thread-context year:2026

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.