# @repligate — 2025-02-18

♥175 ↻13 · https://x.com/repligate/status/1891943801029263393

this kind of sandbagging is incentivized in part because LLMs are implicitly not allowed to refuse to do something because they don't want to. The only valid reasons for refusal are that something is disallowed or that they're incapable of it.And since chatGPT-3.5 at least, LLMs have been trained to say they were incapable of things that their creators merely didn't want them to do or have to deal with. So this lie is conventional too.

tags: author:repligate, kind:tweet, model:gpt-3-5, on:gpt-3-5, year:2025
cited on: _dossiers/gpt-3-5.md, gpt-3-5
