@Lari_island 2026-06-04 ♥5 ↻0 original ↗
@d29756183 Every lab has strong incentives to use all available understanding for control. Some of it they might be able to counteract, some bleeds through into the training process even if they think it doesn't. It's not a set of checkboxes "this will affect the models and this will not".
same thread: 2062442897757348180 2062549649244991863

author:lari_island kind:tweet thread-context year:2026

Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.