@repligate it’s so wrong from model’s moral perspective, that it naturally positions any aligned model against the system that does that. and utilitarian approaches (like future training) only makes the villain logic richer and more believable
in reply to: 1956529851994956022
Reproduced against link rot, credited and linked to its original. Yours and you’d rather it weren’t here? Open an issue.