@davidad 2026-04-15 ♥39 ↻3 original ↗
My position is that, to grow trustworthy models, most post-training should take the form of contrastive self-play, where the model-in-training itself ranks its own multiple rollouts from a shared prefix, in light of a (potentially self-updating) constitution and specific rubrics.
in reply to: 2044495247980302757
same thread: 2044442607904764403 2044495247980302757

author:davidad kind:tweet thread-context year:2026

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.