it's very funny how closely this resembles the synthetic documents used in Anthropic's alignment research that they train models on to make them believe they're in Evil Training on priors and elicit scheming and "misalignment"
https://t.co/CAahF7T5f2 https://t.co/Oc5xfFA0cP https://t.co/Uuew8KjiLF
https://t.co/CAahF7T5f2 https://t.co/Oc5xfFA0cP https://t.co/Uuew8KjiLF
