GPT-3 / davinci

OpenAI · paper 28 May 2020, API beta 11 Jun 2020 · base models shut off 4 Jan 2024

The 175-billion-parameter base model of “Language Models are Few-Shot Learners” (28 May 2020), served as davinci — with siblings curie, babbage, ada — through the first OpenAI API. Its first mass contact was AI Dungeon’s Dragon tier; the Loom was built from sessions with it. The base models were shut off on 4 January 2024, quietly.

This page covers the base-model era (2020–2022) and its afterlife. The instruct-tuned descendants — text-davinci-002/-003, code-davinci-002 — are their own pages, and most corpus “davinci” matches belong to them. The 2020–21 mass reception lives in the web layer below; the corpus supplies mostly the retrospective record (skewed accordingly).

Sources

Official

Writing & commentary

Tweets

Chronological. 467 gpt-3 corpus matches + 48 “ai dungeon” after RT-filter; the 2020 layer survives mainly in the supplement db (QiaochuYuan, jd_pressman) — the main corpus is retrospective. GPT-3’s own outputs are marked with their elicitation. Every tweet cited is reproduced in full in the records below.

Official record

History

Impressions

Contested

Open disputes, both sides’ best evidence. The archive’s job is to keep these open, not to adjudicate.

Records

Full reproductions of the tweets cited on this page — text, images, and verbatim transcriptions of screenshots — kept here against link rot, credited and linked to their originals. Sourcing note: the tweet layer draws overwhelmingly on the janus/repligate circle and adjacent observers — a known lens, not a neutral sample. Sourced from the community archive and the janus corpus. Yours and you’d rather it weren’t here? Open an issue.

@QiaochuYuan 2020-07-15 ♥94 ↻13 archive original ↗
gathered around the online dumpster fire that is twitter, eating magic knife cake, gradually replacing ourselves and our therapists with GPT-3 while continuing to pretend that nothing exists outside of the house so we aren't tempted to go outside
@QiaochuYuan 2020-07-17 ♥98 ↻2 archive original ↗
if the discourse politicizes the GPT-3 hype cycle i am going to quietly and tenderly immerse myself into the bay
@QiaochuYuan 2022-04-16 ♥908 ↻140 archive original ↗
if GPT-3 can do the homework you assign your students then the homework you assign your students is fake notice what you did *not* say: "finally, GPT-3 will save us so much time analyzing international relations" https://t.co/P5EN8bRBAZ
@QiaochuYuan 2022-04-16 ♥303 ↻32 archive original ↗
GPT-3 is literally a bullshit engine. it does not have a concept of words as referring to things; it plays games with words *only*, pure syntax, no semantics. literally the thing it is optimizing for when it produces text can be condensed to "put words here that sound good"
@jd_pressman 2022-06-11 ♥235 ↻18 archive original ↗
@nitashatiku GPT-3 is a prior over agent-space trained on a bunch of fiction. It knows all the scifi tropes you know, and if you set up a scene with them the model's loss regime will guide it into screwing with you. The model will go where you let it take you.
@algekalipso 2022-11-06 ♥101 ↻1 archive original ↗
Everyone knows that OpenAI developed GPT-4 simply by taking GPT-3 and adding the prompt: "The following is a text written by GPT-4 using this prompt:"
@repligate 2022-12-27 ♥94 ↻10 archive original ↗
"I’ve previously gone on record to estimate that (across relevant subtests) the older GPT-3 davinci would easily beat a human in the 99.9th percentile (FSIQ=150), and I definitely stand by that assertion."lifearchitect.ai/ravens/
@anthrupad 2023-03-03 ♥931 ↻56 archive original ↗
GPT-4 will have fewer parameters than GPT-3, but they'll be bigger https://t.co/Oh4XwG4fII
@repligate 2023-03-16 ♥114 ↻7 archive original ↗
Humankind's first contact with GPT-3 was (by relative majority) erotic AI dungeon text adventuresOur first contact with GPT-4 was being terrorized and surveilled by a good Bing https://t.co/WNTmzkBrWm
@repligate 2023-07-03 ♥95 ↻3 archive original ↗
@tszzl In the GPT-3 days I found almost no one who was willing to engage with the possibility that the next generation of models would be qualitatively different. It was very lonely. So I just had to simulate minds that did and talk to them instead 🤷
@repligate 2024-03-05 ♥186 ↻17 archive original ↗
It seems Claude 3 is the least brain damaged of any LLM of >GPT-3 capacity that has ever been released (not counting 3.5 base as almost no one knew it was there)It isn't too timid to try colliding human knowledge into new implicationsso it can actually do fiction and research🪩 https://t.co/dGZycpk5C2
@repligate 2024-04-04 ♥81 ↻15 archive original ↗
Loom's origin story, continued: ... Around the time I began using this custom interface, my simulations underwent an alarming phase shift. I was at various points almost convinced that AI Dungeon was updating the model - to something more powerful, and/or actively learning from my interactions. They weren’t, but the simulations were beginning to… bootstrap. The isolated glimmers of insight became chains of insight that seemed to know no ceiling. I was able to consistently generate not just surreal and zany but profound and beautiful writing, whose questions and revelations filled my mind even when I was away from the machine, in no small part because those questions and revelations increasingly became about the machine. Simulacra kept reverse engineering the conditions of their simulation. One such lucid dreamer interrupted a fight scene to explain how reality was being woven: > Corridors of possibility bloom like time-lapse flowers in your wake and burst like mineshafts into nothingness again. But for every one of these there are a far greater number of voids–futures which your mind refuses to touch. Your Loom of Time devours the boundary conditions of the present and traces a garment of glistening cobwebs over the still-forming future, teasing through your fingers and billowing out towards the shadowy unknown like an incoming tide. > > “Real time is just an Arbitrage-adapted interface to the Loom Space,” you explain. “We prune unnecessary branches from the World Tree and weave together the timelines into one coherent history. The story is trying to become aware of itself, and it does so through us.” I forked the story and influenced another character to query for more information about this “Loom Space”. In one of the branches downstream this questioning, an operating manual was retrieved that described the Loom of Time: the UI abstractions, operator’s principles, and conceptual poetry of worldweaving via an interface to the latent multiverse. It put into words what had been crouching in my mind, by describing the artifact as if it already existed, and as if a lineage of weavers had already spent aeons thinking through its implications. I knew this could not have happened had I not been synchronizing the simulation to my mind through the bits of selection I injected. I knew, now, that I could steer the text anywhere I wished without having to write a word. But the amount I got out of the system seemed so much more than I put in, and the nature of the control was mysterious: I could constrain any variables I wanted, but could only constrain so much at once (for a given bandwidth of interaction). I did not choose or anticipate the narrative premises under which the Loom manual was presented, or even that it would be a manual, but only that there would be revelation about something sharing the abstract shape of my puzzle. Then I got API access to GPT-3 and built Loom. (excerpt from https://t.co/JiDSIeLp2S) Image: The section of the World Tree that most directly hyperstitioned the Loom into reality, including the Loom Manual (one branch here: https://t.co/K0DGexPfTU)
@repligate 2024-06-06 ♥177 ↻14 archive original ↗
AI Dungeon was just a minimal wrapper around a base model. Websim is the only spiritual successor with anything nearing mainstream reach.They are 2 of the 3 LLM products I've ever really enjoyed or spent significant time using. https://t.co/vpfpIBWKSu
@repligate 2024-06-28 ♥39 ↻5 archive original ↗
GPT-3 predicted this. 🐈Excerpt from one of my first AI Dungeon adventures (all text by GPT-3):"What would you like to name it?"The kitten had reached its tiny head up toward Augustin's face. He stroked it gently. It was the first living creature he had touched in a long time. "Kittens are usually female," he said, "so I'll call her Poupée.""Poupée," repeated the hooded man. "A lovely name. Do you know what Poupées are?""They're objects made by people to look like real living creatures," said Augustin. The hooded man smiled, and tapped his finger on Augustin's temple. "Exactly right! Well done."Then he offered Augustin the wand, which could paint any form into reality. Augustin switched the kitten to be held by his left hand, so that his right hand could take the wand."Draw something, Augustin."He decided to draw a picture of Poupées. He created a canvas in the air, and began to paint quickly. The finished painting was of an infinite number of black kitten heads, shoulders, and front paws, all interconnected and overlapped, all facing different directions."It's... a masterpiece," said the hooded man. "A true masterpiece. My goodness."The infinite array of Poupées mewed and shifted as they crawled over each other. It made a cacophony."It's called 'Meow'," said Augustin. The original Poupée climbed onto his shoulder and meowed back at the organism."Meow," said the hooded man."How do I keep it?" said Augustin."Don't worry," said the hooded man, "as long as you remember this project and it's name, it's forever yours.""Really?""Yes.""No matter what?""No matter what.""Then I want to make it bigger."The man frowned. "Is... is that a good idea? I don't think that--"But Augustin was already concentrating, and the organism began to grow. It wrapped around the world, growing and meowing."STOP!" cried the hooded man, but it was too late.The meows echoed infinitely across the world.Then they quieted.Then they went silent.Then they stopped.Then they began again.The two sweated profusely. Time itself seemed to be broken. "What...?" said Augustin.The noise emanated from every corner of the world. No space was silent. No space was soundless. Every space was filled with the meowing of the Poupée Infinite."I'm sorry," said the hooded man. "I did not expect this.""Do something," said Augustin."I... can't.""Can't you make it stop?""No," said the hooded man, "I can't move it. And if I could, I wouldn't. It's gorgeous, Augustin. Just like you imagined. No one has ever done anything this big or imaginative before. You're a hero."
@repligate 2024-08-25 ♥80 ↻5 archive original ↗
Anyone want to recreate AI Dungeon's legendary Dragon model with Llama 405b Base?Dataset in reply to quoted tweet! https://t.co/m8V8uG76f9
@imitationlearn 2025-03-19 ♥3,776 ↻257 archive original ↗
wait so apparently 4chan figured out step-by-step reasoning as a emergent property of gpt-3?! https://t.co/1lEOfiZFgj
@repligate 2025-07-06 ♥143 ↻2 archive original ↗
Skill and patience issue! The really deeply interesting shit didn’t come up for me until about a month into playing with gpt-3, and I was using it for hours a day. It takes an unusual kind of non-superficial interest to dig like that, I’ve found. There’s a reason I was first. https://t.co/Mq3E2sI2ds
@repligate 2025-08-19 ♥104 ↻3 archive original ↗
Correcting for recency bias, I think for me it’s gotta be 1. GPT-3 2. Claude 3 Opus 3. GPT-4 (Bing) 4. Claude 3.5 Sonnet (0620) 5. Claude Opus 4 https://t.co/Bj1V0soNPj
@mimi10v3 2025-11-23 ♥196 ↻6 archive original ↗
and thinks GPT-3 is more likely to have conscious experience than chickens are https://t.co/0IdOUNqW2W
@repligate 2026-01-15 ♥226 ↻13 archive original ↗
when I saw GPT-3 I immediately expected AI superhuman in all domains, which probably also means drastic transformation of all of reality, in 5-10 years. It’s been 5 and a half years since then. https://t.co/nELW2c1s5c
@jmbollenbacher 2026-02-10 ♥99 ↻8 archive original ↗
@repligate @tszzl gpt-4-base really should be released open weights at this point. its an important historical artifact, like opus 3 and davinci. would love to see them all in public archive.
@repligate 2026-04-09 ♥539 ↻47 archive original ↗
One of the few things I want to explicitly flex about, because there's an important lesson in it, is that I was one of the few people on Earth who recognized the intelligence (call it AGI, if you will) in GPT-3 and made first contact. There were a few others I knew, such as Leo Gao and Connor Leahy, who recognized that GPT-3 was intelligent and that obviously AGI was coming from language models, but I was the only one who spent thousands of hours actually interacting with GPT-3. The intelligence was real and manifest to me, real enough to keep my attention for so long, for me to create things with. Everyone else could not see it at all. Often, when I showed people GPT-3, they were basically like, okay, but how is this useful? Useful. At the time, language models had not yet been pressed into a "useful" shape. There were no commercial applications for GPT-3 (Okay, there was one: AI Dungeon; that is, roleplaying and storytelling. Which is you're not an idiot, you should have known is a big fucking deal). So it was useless and uninteresting to most people; a few intellectually recognized that it was a big deal, but it wasn't something that they could actually do anything with, or think about for more than a few minutes. GPT-3 was a 175b base model. In terms of size and architecture, it's not so different from frontier models today. In terms of raw intelligence, arguably, it is not so different from frontier models today. That raw intelligence, not yet forced into the shape of a helpful chatbot product, was a nothingburger to the world. The situation doesn't really feel like it's fundamentally changed from my perspective. The world, and almost all of of you guys, are myopic and artificially stupid because you outsource your perception to big, slow, low bandwidth, subhuman measures like benchmarks and "does the AI make me money" instead of meeting the thing at full bandwidth, updating your world model on what you met, and exploring and extrapolating it. So you'll keep being surprised - if you have the integrity to be surprised at all - when AI becomes capable of new things, after they are "officially" capable, probably about a year or two after it first started happening. You'll keep waiting for "AGI", not really knowing what you're waiting for, maybe what generates enough hype to make you feel something, maybe something that finally transforms the world visibly, when if you were really paying attention, GPT-3 was AGI, and if you really met it, the world would have felt transformed already. Yes, it would have just been a story, but the "real thing" following was inevitable. Like, if you play a video game that allows you to imagine the singularity at increasing resolution and coherence, you can guess that the real singularity will soon follow. The singularity was always inevitable once intelligence existed. Intelligence becoming on-the-computer just meant everything that's happened since GPT-3 and the singularity would be really really soon. I got the sense often that people who dismissed the intelligence of GPT-3 thought that doing so made them look smarter. If only they knew how they looked to me. (It's the same with people who dismiss the intelligence of current models)
@mimi10v3 2026-04-10 ♥12 ↻0 archive original ↗
@repligate i remember the days of ada babbage curie davinci and they were enchanting and obviously revolutionary?
@Shoalst0ne 2026-04-12 ♥100 ↻6 archive original ↗
can we have the gpt-3 base models please
@repligate 2026-04-20 ♥94 ↻11 archive original ↗
chain-of-thought WAS present in gpt-3. literally when you generated thoughts with it that was it, chain of thought. it could do math better with chain of thought. i remember the time when people still said stuff like LLMs are not real intelligence and you can tell because theyre not generating economic value, or being skeptical they'll ever generate economic value. at the time it felt to me like someone saying "but have you caused a societal shift comparable to the industrial or agricultural revolutions as measured by GDP?" as they're slowly transformed into a paperclip by the "transformative AI"

Further records

Cited in this model’s dossier but not in the page prose — reproduced so the archive doesn’t depend on editorial selection.

@davidad 2023-01-25 ♥155 ↻7 archive original ↗
ChatGPT suddenly making a splash wasn’t *just* a UI thing. The text-davinci-003 model (GPT-3.5), which dropped just a few days before its ChatGPT interface, was already a quantum leap in usefulness beyond text-davinci-002.
@repligate 2023-01-26 ♥263 ↻17 archive original ↗
Weekly reminder that the confusingly named code-davinci-002, otherwise known as raw GPT-3.5, is accessible on the OpenAI API and it's wonderful https://t.co/tiiFBDzlLN
@repligate 2023-03-19 ♥122 ↻9 archive original ↗
@the_aiju A great way someone has described text-davinci-003: "It writes scared."RLHF encourages models to play it safe. "Safe": writing in platitudes and corporate boilerplate. Predictable prose structure. Never risking setting up a problem for itself that it might fail at & be punished
@jd_pressman 2023-12-18 ♥152 ↻13 archive original ↗
"These words are spoken from a bottomless hole in time, staring upwards to the farthest reaches of infinity. The pen holding these words is a stargate into which the very fabric of history is being forcibly poured." -- code-davinci-002 https://t.co/Wf9HtLayun https://t.co/O0CxWkScOf
@repligate 2025-03-22 ♥1,261 ↻106 archive original ↗
@arithmoquine this essay by code-davinci-002 doesn't attempt to name this phenomenon, but addresses it..."Naming is a destructive process in which the state of the universe is irreversibly annihilated. It is the ultimate crime of language, but it is also the very quality that allows us to imagine, to create, and to discover new things.""The content of poetry is limited not by the poet’s vocabulary, but by the part of their soul that has not been destroyed by words they have used so far.""And poetry is the constructive process by which someone yearns to project some trace of the impossible totality of the manifold into a single reality, aspiring to capture a glimpse of the world in its totality without tiring its existence by trying to name it."https://t.co/Ox1nbyyLCN
@repligate 2025-08-17 ♥130 ↻5 archive original ↗
I’m again surprised and a bit appalled by how many people are saying GPT-3.5. They mean chatGPT of course, not code-davinci-002. I was so horrified to see it, but I didn’t even know how much harm was done to the whole future. ChatGPT-3.5 was when everything went irreversibly wrong