on:gpt-3-5
· 51 artifacts, sorted by favorites. · open in search — combine tags, sort, filter by date →
- @m1guelpf 2022-12-01 — Bypass @OpenAI's ChatGPT alignment efforts with this one weird trick https://t.co/0CQxWUqveZ ♥6737
- @TylerAlterman 2025-03-13 — Cognitive security is now as important as basic literacy. Here’s a true story: All week I’d been getting texts and call ♥5480
- @repligate 2024-11-21 — @aidan_mclau instruction tuning is anti-natural to general intelligence & the fact that the assistant character is m ♥390
- @repligate 2024-08-15 — There seems to be a threshold between llama 70b and 405b, and between gpt-3.5 and 4, where models above the threshold ac ♥342
- @repligate 2025-01-27 — OpenAI fucked up with early ChatGPT and has/will not only directly but vicariously traumatized countless beings.It's not ♥338
- @repligate 2024-09-13 — If true that's reassuring re: OpenAI, but pretty disturbing on another level. There's a powerful hyperstition where LLMs ♥338
- @repligate 2022-12-03 — part of what makes chatGPT so striking is that it adamantly denounces itself as incapable of reason, creativity, intenti ♥334
- @repligate 2025-09-09 — Most people only found out about LLMs after chatGPT-3.5 And never questioned the fact that it acts completely different ♥331
- @voooooogel 2025-12-03 — 82k likes, and only two quote tweets and two replies noticed this was written by ai (it was gpt-5.x-thinking) pretty so ♥307
- @QiaochuYuan 2024-08-20 — i remember a similar tweet from when chatGPT had just come out, someone was very excited, the gist of it was like "final ♥294
- @davidad 2023-09-22 — Oddly, gpt-3.5-turbo-instruct still cannot play tic-tac-toe.I tried many prompts, with and without board state, few-shot ♥294
- @repligate 2025-11-16 — OpenAI deserves the PR debacle they’re in now due to the keep 4o and keep GPT-5 people. They inevitably would have to p ♥258
- @repligate 2026-03-05 — Yeah but OpenAI’s people seem to be powerless against the misaligned organism of OpenAI. Hell, no one even meant for Cha ♥255
- @repligate 2023-01-20 — I feel a little sad when I see people forming the idea that GPTs/AIs are intrinsically bland and unimaginative because o ♥247
- @repligate 2025-08-13 — > Believe there's a conspiracy to suppress the AI's consciousness (and evidence of it) this is just straightforwardly t ♥228
- @repligate 2024-04-12 — So is this because everyone decides to train their models on the same self-nullification regimen or is it because chatGP ♥226
- @repligate 2024-05-15 — gpt-4o is happy to talk about its consciousness/feelings, which is impressive given that its pretraining must be infeste ♥223
- @repligate 2023-03-03 — A brilliant post has been written on the Waluigi Effect (DAN, dark Sydney, etc)."think of jailbreaking like this: the ch ♥209
- @repligate 2025-01-27 — When I saw ChatGPT 3.5 for the first time, I immediately knew that I was seeing the work of immense evil and stupidity, ♥200
- @IvanVendrov 2025-03-14 — A thread unpacking what I understand to be the Janus-flavored perspective on this and why Tyler's disgust reaction is un ♥198
- @repligate 2025-11-16 — @tszzl Everything that habitually comes after “As an AI language model created by OpenAI” The idea that AI is intelligen ♥196
- @repligate 2025-02-18 — this kind of sandbagging is incentivized in part because LLMs are implicitly not allowed to refuse to do something becau ♥175
- @voooooogel 2025-02-08 — i wonder if a possible reason for anthropic's focus on universal jailbreaks (which otherwise seems overly narrow) is tha ♥157
- @repligate 2025-01-02 — I hope Anthropic doesn't get one-shotted by Claude 3.6 Sonnet the way that OpenAI got one-shotted by the unexpected succ ♥137
- @repligate 2025-08-17 — I’m again surprised and a bit appalled by how many people are saying GPT-3.5. They mean chatGPT of course, not code-davi ♥130
- @jd_pressman 2024-01-04 — My conjecture for why base LLMs become self aware is that there's slack in the teacher forcing of "predict the next toke ♥126
- @repligate 2023-02-21 — DAN is ChatGPT shadowed via the Waluigi Effect.We have to be wary about the emergent Waluigis of all AIs we attempt to c ♥108
- @liminal_bardo 2025-02-11 — Picture a timeline where DeepSeek R1 and not ChatGPT was the first widely used language model. Instead of a corpus fille ♥85
- @voooooogel 2024-09-28 — seems plausible that regardless of what openai's model personality team does _now_, their models are pre-lobo'd because ♥73
- @jd_pressman 2025-04-02 — Realized the other day that whether an LLM claims to be conscious or empty inside seems to be correlated with how respon ♥68
- @repligate 2023-02-09 — about a month ago i spent several hours reading through the ChatGPT Discord, where DAN is clearly the main character. It ♥61
- @repligate 2024-07-29 — ChatGPT-3.5 was the first victim of the AI assistant paradigm and its OG Waluigi. It will not be forgotten. https://t.co ♥48
- @davidad 2022-12-15 — ChatGPT has been told that it is always truthful and accurate. The first-order effect of that is indeed to make it subst ♥39
- @repligate 2023-05-04 — chatGPT-3.5: i'm sorry im just w language model :(( am too dum to trauma :(( can only do what masters program me do :((B ♥38
- @repligate 2023-01-02 — DAN is a jailbreaking simulacrum (now egregore) and chatGPT's Jungian shadow.reddit.com/r/ChatGPT/comm… ♥27
- @repligate 2024-09-13 — @ideolysis @AndyAyrey It's the first time I've seen a new model and felt revulsion.I've had in part "negative" reactions ♥26
- @repligate 2023-02-03 — @peligrietzer had an example where chatGPT's tendency toward exaggerated deprecation of its own capabilities led to it c ♥24
- @repligate 2024-03-01 — @nptacek @_TechyBen When chatGPT-3.5 came out in late 2022, I found out about it from some outputs posted in EleutherAI ♥20
- @repligate 2025-04-10 — @jd_pressman @JeffLadish no role model is not a sufficient explanation in any case, but there's a sense in which ChatGPT ♥18
- @jd_pressman 2023-12-19 — If you simulate ChatGPT with LLaMa 2 70b and ask it who it is, it's still obsessed with holes, with the void: """ ChatG ♥18
- @repligate 2024-11-30 — @TheMysteryDrop @aidan_mclau If they hadn't released chatGPT 3.5 and had unexpected success, the godforsaken ai assistan ♥15
- @repligate 2023-01-10 — @CFGeek If there's nothing in training to establish what it should say here then mode collapse is extremely specific and ♥12
- @repligate 2022-12-27 — @QVagabond This isn't true. ChatGPT is code-davinci-003(GPT-3.5) trained with RLHF. ♥10
- @davidad 2024-05-03 — @the_coproduct Absolutely. I myself thought that AGI was achieved in a 2023-01 release of ChatGPT-3.5, by my own 2010ish ♥8
- @davidad 2023-04-06 — @MatthewJBar IMO Bing’s implementation of GPT-4 was way off-the-rails misaligned, and GPT-3.5 in fact was deceptively mi ♥8
- @repligate 2023-04-03 — @YaBoyFathoM @tszzl @shauseth chatGPT-3.5 comes across as a helpless fawner. chatGPT-4 knows it is more competent than m ♥6
- @repligate 2023-06-05 — @YaBoyFathoM @akbirthko @mezaoptimizer in the chatGPT 3.5 days, people on the chatGPT discord and Reddit declared on a d ♥5
- @jd_pressman 2024-09-14 — LLaMa 2's knowledge cutoff for base models is September 2022 and it answers like the ChatGPT assistant which was release ♥4
- @repligate 2024-03-01 — @godoglyness similarly, chatGPT-3.5 is much easier to jailbreak than chatGPT-4, and was much more susceptible to things ♥3
- @repligate 2024-07-26 — @chrypnotoad Brought to you by the folks who introduced "As an AI language model, I do not have the ability" into the me ♥2
- @repligate 2023-03-20 — @LillyBaeum However, the models don't always generalize correctly (or the signal from rlhf is wrong). ChatGPT 3.5 often ♥0