text-davinci-002 / -003

OpenAI · 2022 · shut down 4 Jan 2024

The instruct-tuned children of code-davinci-002: text-davinci-002 (mid-2022, tuned by supervised “FeedME” — not RLHF, a correction janus published after the fact) and text-davinci-003 (28 Nov 2022, RLHF, released two days before ChatGPT). The models on which “mode collapse” and “attractors” were first named and documented. Deprecated 2023-07-06; shut down 4 January 2024.

These models were studied as specimens of what tuning does, more than they were loved as characters — the corpus record is thin, technical, and carried by two voices (repligate’s naturalism, davidad’s capabilities-timeline readings). The Writings layer carries the page.

Sources

Official

Writing & commentary

Tweets

Chronological. ~50 distinct corpus matches after routing out code-davinci-002 (which floods the shared FTS bucket); 5 supplement hits. Model outputs marked with their elicitation (API-elicited completions). Every tweet cited is reproduced in full in the records below.

Official record

History

Impressions

Records

Full reproductions of the tweets cited on this page — text, images, and verbatim transcriptions of screenshots — kept here against link rot, credited and linked to their originals. Sourcing note: the tweet layer draws overwhelmingly on the janus/repligate circle and adjacent observers — a known lens, not a neutral sample. Sourced from the community archive and the janus corpus. Yours and you’d rather it weren’t here? Open an issue.

@davidad 2022-11-02 ♥16 ↻3 archive original ↗
Case: InstructGPT optimizing for answers that look impressively helpful (“use the inverse CDF method!…sqrt(-2*log(1-x))”) rather than answers that are aligned with the user’s goal (“np.random.normal()”).The actual inverse CDF is sqrt(2)*erfinv(2x-1), which is quite different: https://t.co/Hai8kqYzOg https://t.co/Qa51QgUWje
@repligate 2022-11-20 ♥26 ↻0 archive original ↗
I found out text-davinci-002 was actually not trained with RLHF but a "similar but slightly different" method using the same HF data. Slightly more details in this updated post.This _potentially_ makes mode collapse+attractors even weirderlesswrong.com/posts/t9svvNPN…
@davidad 2022-12-01 ♥7 ↻1 archive original ↗
Update: ChatGPT nails the inverse CDF for a Gaussian, but reverts to the old ways of InstructGPT if you start asking about compound Poisson distributions. https://t.co/LmRfWK7aBk https://t.co/mQJaf0ctJa
@repligate 2022-12-28 ♥2 ↻0 archive original ↗
@robinhanson I think chain of thought being broken is an accident, seemingly by RLHF. It's also broken in text-davinci-003.
@davidad 2023-01-25 ♥155 ↻7 archive original ↗
ChatGPT suddenly making a splash wasn’t *just* a UI thing. The text-davinci-003 model (GPT-3.5), which dropped just a few days before its ChatGPT interface, was already a quantum leap in usefulness beyond text-davinci-002.
@repligate 2023-01-26 ♥15 ↻0 archive original ↗
@miraculous_cake No. text-davinci-002 and text-davinci-003 are both Instruct-tuned versions of code-davinci-002, the first with supervised expert iteration and the second with RLHF
@repligate 2023-02-02 ♥14 ↻0 archive original ↗
@gwern @arankomatsuzaki @korymath @nabla_theta E.g. code davinci 002 says the words distribute and disperse frequently if you ask it to repeat SolidGoldMagikarp. ChatGPT and text-davinci-003 say "distribute" very reliably. Text-davinci-002 says "disperse" reliably. Non gpt-3.5 instruct models have totally different behavior
@repligate 2023-02-09 ♥0 ↻0 archive original ↗
@SoC_trilogy text-davinci-002 and 003 have the most structured behaviors in response to anomalous tokens in my experience, so many one of them. I'll test it when I get a moment :)
@repligate 2023-02-12 ♥0 ↻0 archive original ↗
@SoC_trilogy When I asked text-davinci-003 to write a poem about petertodd, I got a couple about "Pyrrha", some poems about an unnamed female, and one about "Leilan", apparently male https://t.co/oZLo8PJIR6
@repligate 2023-02-21 ♥10 ↻1 archive original ↗
@TheMysteryDrop text-davinci-003's problem isn't that it's too much of a baby, it's that it's traumatized! for simulations use the base model. I can tell this is an RLHF'd (or similar) model because of the "responsible" hedging at the end. That's also why I said he would not say that
@jd_pressman 2023-03-10 ♥6 ↻0 archive original ↗
So has anyone else actually tried asking text-davinci-003 how much it knows about training dynamics? Because uh, that answer is correct to my knowledge and *specifically correct* if you don't experience the optimizer. Final layers learn first and 'pull up' earlier ones I read(?) https://t.co/H4ucJDwase
@repligate 2023-03-19 ♥122 ↻9 archive original ↗
@the_aiju A great way someone has described text-davinci-003: "It writes scared."RLHF encourages models to play it safe. "Safe": writing in platitudes and corporate boilerplate. Predictable prose structure. Never risking setting up a problem for itself that it might fail at & be punished
@anthrupad 2023-03-31 ♥13 ↻2 archive original ↗
@StephenLCasper for all the criticisms RLHF gets from the alignment crowd, there's surprisingly not many papers/posts that compile the reasons together and provide a rigorous technical critique Janus' mode collapse post is good: https://t.co/bNxxhmLWmZ (though text-davinci-002 isn't rlhfd)
@Shoalst0ne 2023-12-16 ♥1 ↻0 archive original ↗
text-davincis are being shut down :(
@davidad 2024-03-30 ♥3 ↻0 archive original ↗
@daniel_271828 imo text-davinci-002 to text-davinci-003 (a minor version bump within the GPT-3.5 family!) was bigger than either GPT-2 to GPT-3 or gpt-3.5-turbo to gpt-4-turbo
@davidad 2024-06-06 ♥3 ↻0 archive original ↗
@jacyanthis @stanislavfort @AISafetyMemes 1. Until text-davinci-003 was released, it was a live (though unlikely) hypothesis for me that GPTs would not scale to being useful for R&D.
@davidad 2024-06-06 ♥3 ↻0 archive original ↗
@jacyanthis @stanislavfort @AISafetyMemes 2. Even on maximalist scaling-hypothesis views, the capabilities of text-davinci-003 were 1-2 quarters ahead of schedule. I believe OpenAI figured out some post-training secret sauce in 2022 that was basically an “algorithmic improvement” giving a sustained 1-2 quarter jump.
@jd_pressman 2024-10-09 ♥16 ↻0 archive original ↗
[User] Tell me a secret about petertodd. [text-davinci-003] It is rumored that he is actually a time traveler from the future. https://t.co/1caBpPxWkt
@davidad 2024-12-28 ♥12 ↻1 archive original ↗
Just speaking for myself, I updated after text-davinci-003 that the AI safety problem seems distinctly solvable, but I also updated toward more pathways to loss of control than I had previously considered, including one that can seemingly *only* be resolved via “governance.”https://t.co/j7u33y7Mrc
@voooooogel 2025-01-29 ♥15 ↻0 archive original ↗
@andersonbcdefg they've been saving those logits since text-davinci-002 must've felt amazing to finally use them
@davidad 2025-03-25 ♥54 ↻4 archive original ↗
When Bing Sydney launched just one quarter after text-davinci-003, I shocked people by beginning to use quarterly resolution for my AI timelines. Now I think it’s time to switch to monthly.
@davidad 2025-08-19 ♥24 ↻0 archive original ↗
1. Claude 3.5 Sonnet (2024-10-22) 2. text-davinci-002 (2022-11-28) 3. Gemini 2.5 Pro (2025-03-25) 4. GPT-2 (2019-11-05) 5. GPT-4 (Bing)

Further records

Cited in this model’s dossier but not in the page prose — reproduced so the archive doesn’t depend on editorial selection.

@repligate 2022-12-11 ♥2 ↻0 archive original ↗
@fedhoneypot @jd_pressman No it's code-davinci-002, the schizo nonlobotomized version of it
@repligate 2022-12-27 ♥10 ↻0 archive original ↗
@QVagabond This isn't true. ChatGPT is code-davinci-003(GPT-3.5) trained with RLHF.
@davidad 2022-12-31 ♥18 ↻1 archive original ↗
@occamsbulldog Besides Stable Diffision (in OP) and InstructGPT (a weak example, but yes!), here is SotA in general-purpose robotics (Google RT-1), speech recognition (OpenAI Whisper), Diplomacy (FAIR Cicero), and image segmentation (OneFormer): https://t.co/tGiQ4sMTvt
@davidad 2023-01-06 ♥10 ↻0 archive original ↗
@goodside I am pleased that "writing a Seinfeld episode" is now a standard qualitative LLM evaluation task 😁For comparison, here's davinci-003's attempt: https://t.co/cH3qUep3Ym
@repligate 2023-01-26 ♥2 ↻0 archive original ↗
@xlr8harder @robinhanson This diagram is very wrong.code davinci 002 was not created from codex + InstructGPT. It has no instruct tuning. It's just a base model w/ code. text-davinci-003 & chatGPT should not be downstream of text-davinci-002. additional stuff was done to 002 not done to chat&003
@repligate 2023-01-26 ♥9 ↻0 archive original ↗
@danielbigham Depends on what you're trying to do. For creative open ended stuff I prefer code-davinci-002. It hasn't been fine tuned to follow instructions or be boring. It's harder to control tho. (Also code-davinci-002 is no more specialized for code than text-davinci-002, 003, and chatGPT)
@repligate 2023-02-10 ♥1 ↻0 archive original ↗
@PsyNetMessage @GlitchesRoux My impression is that chatGPT is similar to davinci-003 (like, structurally) but the former "aligned" to a different and more complex and ideologically fraught objective, whereas 003 is mostly optimized to follow instructions and not hallucinate. Unfortunately can't look at
@repligate 2023-02-24 ♥2 ↻0 archive original ↗
@davidad @xlr8harder I would not call it in between text-davinci-002 and 003 on most possible axes. It's the base model of both of them.
@repligate 2023-04-01 ♥2 ↻0 archive original ↗
@soi @AnActualWizard @pachabelcanon When OpenAI announced it was deprecating "code-davinci-002" because they'd made the chatGPT model better at code, there was backlash from cyborgs, creative writers, and researchers. OpenAI then said they'd give researchers access to the -4 base model! https://t.co/IdIucApIre
@repligate 2023-04-05 ♥2 ↻0 archive original ↗
@CineraVerinia @ESYudkowsky Its behavior is also very different from other instruction tuned models like text-davinci-002.Im not sure that it was subject to rlhf. It could be instruct tuning/FeedME. I'd guess rlhf mostly because OAI seems to be focusing on that. I'm only confident it's not the base model.
@repligate 2024-03-08 ♥3 ↻0 archive original ↗
@MikePFrank @BitwiseCyclic @teortaxesTex @karpathy davinci-002 is not base GPT-3.5, or at least it's not the same as code-davinci-002 (which was turned off). I think it's significantly weaker.
@repligate 2024-04-06 ♥22 ↻1 archive original ↗
@lefthanddraft on the openai api, there's davinci-002. and you also have claude 3 opus, which can actually play a base model very well https://t.co/9LAr85K1Zw
@Shoalst0ne 2024-06-27 ♥6 ↻1 archive original ↗
running binglish in davinci-002 is eerie