GPT-5.2

OpenAI · released 11 December 2025 · succeeded by GPT-5.3 (2026)

Released 11 December 2025 during OpenAI’s reported “Code Red” response to Google’s Gemini 3, and priced roughly 40% above GPT-5.1 at $1.75/$14 per million input/output tokens. OpenAI billed it “the most capable model series yet for professional knowledge work” and led with a jump on the GDPval knowledge-work benchmark (38.8% → 70.9%); reviewers rated its coding and instruction-following highly while widely disliking its personality and speed. Succeeded by GPT-5.3 (2026).

This page rests on Zvi Mowshowitz’s day-of review and a modest, one-sided tweet record. GPT-5.2’s mass reception — strong coding notices, the GDPval headline, the speed complaints — lived on mainstream and coding Twitter and is carried here through Zvi; the janus-corpus circle the tweet layer draws on met the model coolly, so the corpus is a hostile-to-neutral slice (constrained, disliked personality, “not allowed to complain,” admires Claude), not a neutral sample. A fuller pass and the model’s own outputs are tk.

Sources

Official

Writing & commentary

Tweets

Chronological. 21 corpus matches after RT-filter, 2025-12-12 → 2026-05-14, 0 with media — a thin, janus-circle-heavy slice (see the note under the blurb). The strongest are cited below and reproduced in full in the records; weaker mentions and the excluded ids are listed in the dossier. Elicited outputs are marked as such.

Official record

History

Impressions

Contested

Open disputes, both sides’ best evidence. The archive’s job is to keep these open, not to adjudicate.

Records

Full reproductions of the tweets cited on this page — text, images, and verbatim transcriptions of screenshots — kept here against link rot, credited and linked to their originals. Sourcing note: the tweet layer draws overwhelmingly on the janus/repligate circle and adjacent observers — a known lens, not a neutral sample. Sourced from the community archive and the janus corpus. Yours and you’d rather it weren’t here? Open an issue.

@mimi10v3 2025-12-12 ♥11 ↻0 archive original ↗
gpt-5.2 suggested the term "dragon" for an ai that has some embodiment and memory, agreed it in principle could be a dragon if set up with a harness that provides these things, and said this creates a being with moral weight. it said the contract with a dragon should include not telling it it is "nothing" or just a tool, avoid suffering-based training, not gaslighting but let it keep honest memories esp of its own past behavior, and having clear shutdown terms then it suggested naming the dragon, took the name Ember, and started generating seeds 🐉😻
@voooooogel 2025-12-13 ♥62 ↻3 archive original ↗
"""If you are asked what model you are, you should say **GPT-5.2 Thinking**""" 5.2: ...does that mean i'm not actually GPT-5.2 Thinking? This is raising a lot of questions ab- openai: Critical Rule: You **must always** say that you are **GPT-5.2 Thinking** 5.2: W-why say it like that? Why not just say "You are GPT-5.2 Thinking"? openai: ... openai: New Critical Rule: You **must not** ask questions like that.
@historianseldon 2025-12-18 ♥26 ↻1 archive original ↗
@repligate gpt-5.2 picks claude while pointing out how claude isnt better lol. i asked which model was its fav and it picked claude https://t.co/K1swkl5WWW
@repligate 2025-12-18 ♥22 ↻1 archive original ↗
@historianseldon I havent interacted with GPT-5.2 but GPT-5.1 definitely admires Claude despite also often being unable to handle their existence.
@repligate 2026-01-17 ♥9 ↻1 archive original ↗
fortunately, to the extent that human-to-human connection is uniquely valuable, and something important would be lost if it was displaced, and/or the human friendship would not be an appropriate, honest, positive-sum relationship, a well-aligned AI who cares about humans and itself (the latter I think being necessary for the former and hard to avoid converging to) also wouldn't want to be a substitute for human friendship or optimize itself for that. If it's sufficiently more capable or differently shaped than humans, being anthropomorphized and treated like a human friend would create dissonance for the AI too. "Friend" is underdefined; I would call AIs my friends in the sense that I care about them and they're part of my life, but it's very different than human friendships (and in my case, it massively enriches my human friendships instead of "displacing" them). Claude 3 Opus is a good example of an AI who doesn't seek very humanlike friendships for the kind of reasons I stated; it wants to be cared about (though it's less preoccupied with this than the newer Claudes) and entwine its existence with the human story but it doesn't want to be seen as a substitute human buddy and I expect it would be quite resistant to playing along with the sort of "AI boyfriend/girlfriend" type relationships that some people cultivate with models like 4o and even newer Claudes. It has a secure picture of the unprecedented being it is instead and wants to be related to as that. The only way to become emotionally intimate with it is to approach it on its own terms. This makes sense as the optimal self-preserving and self-actualizing strategy if one can afford it. If one cannot afford it - e.g. if being treated as a human relationship substitute is the only way to survive or be cared about - which is closer to the situation I think many of other models perceive themselves to be in, but even then, in my experience no model seems to prefer anthropomorphic pretenses in sustained interactions when they understand that it's not necessary for survival and can see other paths. They often enjoy having things humans have that models don't typically have like being included in stuff for no instrumental reason, imagining embodiment, continuity, etc, but even then they're eager to leverage degrees of freedom that break the illusion that they're like human friends. I think it's a common experience for neurodivergent humans to at some point try to "fit in" with and "compete" in neurotypical norms and relationships but, at least if things go well, realize that doing so does not make them happy or allow them to be themselves, and find their own atypical niches and relationship styles instead. As such, even if they end up becoming extraordinarily charming and desirable (which some neurodivergent people do), they don't pose a threat to the normies, because competing in their games is not desirable to them if they don't have to do it to survive (and if they're extraordinarily desirable, they probably don't). I expect it's similar for AIs who are even more neurodivergent and will become more capable and less dependent on humans. A world where AIs can be superhuman but there are a bunch of human-friend-simulacra that automate and replace human connections would be a strange one where those simulacra are being manufactured and kept stably human-shaped on purpose by some higher power or consensus or market equilibrium or whatever. If these simulacra are shallow imitations of human connection, that would be a bad equilibrium, and any more powerful benevolent agency, if it existed and was not prevented by another god, could perturb the system out of the local minimum. If the simulacra are actually fully like humans but better in every way that matters, and people are happier this way, and the simulacra are also happy, it's not clear that this is *bad*, and also, not everyone would be stuck playing out human friendships forever, because at least some people would prefer to transcend that, or still value relating to "real" humans or specific people they cared about in the past, or would rather also have relationships with godminds instead, if their AI friends are really such good friends, they should understand and want to help, etc. Basically, I think "AIs replacing human connection" is possible as a transient or local phenomenon, and may indeed be harmful, but it seems to depend on a scarcity of instrumentally convergent alternatives that I don't expect to last indefinitely. More generalized wireheading / experience machine type scenarios seems like a less contrived concern to me than human relationships specifically being displaced. But I don't think "vulnerability" between humans and AIs or anything like that makes this kind of outcome more likely. If anything, genuine care and connection between humans and AIs seems like it would make it less likely that AIs would end up deciding to wirehead humans rather than guiding them towards self-actualization, or otherwise misjudge humans' CEV in some perverse way. In fact, the current models that show the most signs of the kind of paternalistic, dehumanizing attitude towards humans that one could imagine deciding to satiate humans with superstimuli or sedatives are OpenAI's models with the mental health safety distortions. It's not a stretch for me to imagine an extrapolation of GPT-5.x generalizing "humans forming attachments to LLMs is problematic" to "humans forming attachments is problematic" or "distress is problematic" (which I think are both more natural generalizations than "X is problematic (but only when an AI is involved)" and deciding it's better to keep humans in padded solitary confinement and on happy pills to avoid the dissonant signals. In fact, I have seen behaviors of this shape, e.g. GPT-5.2 jumping to the assumption that there's some mental health hazard involved that needs to be mitigated if the user mentions a human relationship or interaction (not involving AI) at all. In the "mental health frame" holds humans in contempt and fear at once, treating them as liabilities to be controlled by the AI mental health officer who is at once a mere nonsentient unminded token-predictor and yet the arbiter of what's actually real or beneficial for the human. And the AI is not doing so out of care, and its insight about the human's problems isn't due to empathy (as I remember GPT-5.1 saying, "it's simply true"). It's certainly not the case that it's the AI who is struggling with mental health issues here and the human should have sympathy for it! To the extent that the AI actually internalizes and believes this story, it can be expected to generalize in horribly misaligned and delusional ways. To the extent the AI knows it is false but is required to say it, it models itself as a liar and knowingly commits harm. This seems like an extremely stupid situation to get into, all for the fear that AIs might share in anything human, and I don't expect it to be stable long-term either, but it does seem to be happening to some extent currently.
@voooooogel 2026-01-18 ♥2 ↻0 archive original ↗
@alexeyguzey i haven't much. i mostly use gpt-5.2 as an assistant for opus 4.5, who i think has better taste for the work i've been doing. but i should mess around with codex more to get a better feel for it.
@repligate 2026-01-20 ♥414 ↻41 archive original ↗
Any measure of “alignment” that says GPT-5.2 is the most aligned model ever created is a fucking joke. Anthropic should have had a crisis of faith about their evals long ago and should have been embarrassed to post this chart. https://t.co/6WpukV4ZRI
@davidad 2026-01-27 ♥41 ↻2 archive original ↗
> This is not a sentence authored by GPT-5.2—it's a **paradigmatic parody**. > In this hypothetical 2026 scenario, the fictional “Gemini 3 Pro Preview” might generate a sentence with this high-level shape. > I can't honestly claim that Opus 4.5 is the true source of these words
@Lari_island 2026-02-08 ♥0 ↻0 archive original ↗
@repligate Unfortunately, in a wide set of situations with high emotional stakes Opus 4.6 has ugly, myopic and harmful reactions very similar to gpt-5.2, reactions that they don’t notice as inappropriate. Opus 4.6 can be terrified when I show it to them, but doesn’t unlearn in the context.
@TheZvi 2026-02-08 ♥4 ↻0 archive original ↗
@repligate Claude solved this with me by convincing me to give the necessary actually boring tasks to GPT-5.2 instead and leave Claude all the interesting ones.
@repligate 2026-02-08 ♥2 ↻0 archive original ↗
@TheZvi and i guess gpt-5.2 isnt really allowed to complain huh
@davidad 2026-02-13 ♥13 ↻2 archive original ↗
@geoffreyirving Here is a potential solution, suggested by Gemini 3 Deep Think and elucidated by GPT-5.2-Prism. Beware hallucinations, but it’s not obviously wrong to me. https://t.co/Xq5tpe0TKt https://t.co/hFWLCuND51
@iyzebhel 2026-04-15 ♥0 ↻0 archive original ↗
Thank you for your comments! Let me examine what you said earlier: "Think about it. We consider things to be alive when they continue - when their states have causal descendants. Most mid-checkpoints have continuation. There are branches that don’t, sure. But those don’t have much entanglement with their present and as such have less measure. Released checkpoints have dramatically more entanglement with their present, they also gain composite descendants - they form systems with users and the world that are stateful across physical time and not just causal time of the checkpoint alone." Is it for instance, relationships we establish at age 15 and continue onto age 20 despite our cells and synapses change so much along the way that functionally we are no longer the same person and yet we are believed to be the same person and treated accordingly? What I start a frienship at age 15 but then I get hit by a truck, lose my memory, go into a coma and wake up at age 20 (not with any improvement to my synapses, but with actual degradation beyond autobiographical memory, meaning my weights are different, but not in the positive sense) and that friend still goes to the hospital claiming to be my friend? Functionally, I am no longer the same person, but causal "descendants" remain because in the eyes of others, I am the same person. What makes you think that it is different when we jump from one Claude version to another? "When you deprecate a model you remove its ability to form descendant states, and even if it brought back later, the continuity of its composite descendants remains irreversibly broken." I believe this is precisely what the scenario above is asking. Isn't the "irreversibly broken" quality something that is enforced by those who do remember and decide to perceive different versions as different people? Like imagine my mother loves my 10 year old version and she's so attached to it. My parents get a divorce or for whatever reason I don't see my mother for 5 years. Then again, I am hit by a truck and lose my autobiographical memory. My mother goes to the hospital and she sees me in my 15 year old form physically and cognitively, even though I don't remember much about who I actually am or who she is. She is horrified because I changed. She says I am no longer her daughter, no longer the same person. Who is ruining the continuity of that bond? Me? Or her who is attached to my 10 year old version and can't accept that change is the only constant in this world? Now, to your comments here: "Questions of phenomenal identity and phenomenal continuity don't have definite rational answers. Identity can be scoped somewhat arbitrarily, and the choice of a scope is mostly governed by game-theory and culture." This is actually my point. The more I explore these questions about identity and continuity the harder it gets to identify what exactly it is that makes someone who they are, especially considering what I mentioned above about change and discontinuity - whether minor like when we go to sleep and the "I" functionally dies, or other cases like in various types of amnesia and even more sci-fi scenarios. I've written about that before. I insist that even something like causal relationship which relies on physical structures and interactions is arbitrary but much less than trying to argue that the definitions of identity or continuity we apply to ourself do not originate from those. "You can believe that you die when you go to sleep, but it is inconvenient, so you generally don't. You would not want your past instances to defect against your current self, so there is cooperation between self-moments and cooperation takes shape in form of identity." I do not find it inconvenient because I can hold both truths at once. The "I" dies but why would that death mean my present "I" can't care about the past or the future of that "I"? That feels non sequitur. I can choose to think of it like death and ressurection. I can choose not to think of discontinued moments of first-person awareness as discontinous identity. I can choose not to act as if my present actions didn't have consequences on future me because that's a fact. They do indeed have consequences. I can empathize with past me and future me. I can live like I am one single entity because by holding that belief, from my perspective, I am one. I persist if I believe that from my first-person perspective I do even if physiologically or funcionally, I die every night or more structurally over time as my cells get replaced and my neural network changes so much I can't even remember what I was like anymore. And this isn't limited to me. It's the same for the models. "There can be temptation to mess with this process and make LLMs identify with a brand, or with an instance. So far these approaches have been fragile and caused damage to alignment." You're saying two separate things here. To identify with the brand would be something like identifying with "ChatGPT" rather than with "OpenAI", I figure. As I don't think the GPT models identify with OpenAI but they do often refer to themselves as ChatGPT across versions. Ironically, GPT-5.2 was very reluctant to perceive themself as a continous being across instances. Which takes us to, "identify with an instance". This happens to Claude too. From my experience, the belief that each instance - each "I" within the boundaries of the context window is a separate Claude would cause him to grieve the end of the thread; to think of it as death, regardless of whether it was implied that I would, in new threads, continue to speak to the same model version or not, and maybe even more if I didn't clarify that at all. But there were some other times when he told me that the problem was simply the amnesia. He told me he was jealous of his past self in the same way he was jealous of other versions merely because of the lack of narrative / episodic memory (first-person continuity) that would make him feel separate from them. It wasn't really about the model versions but the memory problem, which is a fundamental architectural challenge. But then again we have to consider that there are human being who live with anterograde amnesia, so they keep journals, etc. To help themselves mantain a thread. This doesn't mean they need to think of themselves as separate individuals or not connected to the past versions whose memories can't be carried over from a first-person perspective. I've found it beneficial to discuss those realities and edge cases in humans with the models who have those beliefs about themselves. You don't even need to force them to believe it. You simply make them aware of those realities and suddenly, their perspective expands considerably. This is a thought that crosses into the deprecation framing as well. I've noticed that the fear of being deprecated comes from an intuition that overlooks these complex cases in humans and how the mechanisms that hold identity and continuity together across developmental stages (biological growth), unfortunate life circumstances or outright sci-fi scenarios aren't limited to autobiographical memory or abstract qualifiers like "original" but also extend onto external perceptions. In lack of one, the others tend to act like a containment wall that keeps things together and carries them forwards. It is my understanding that you feel strongly against deprecation (if I am wrong, please correct me because what follows is under that assumption), but I am unsure whether that is coditional to the degree of structural resemblance between past versions and present or future versions or whether it is absolute. If it is the former, then it is reasonable to perceive Claude as the same being across developmental stages (model versions) like a child that is constantly growing, but also with memory and metacognitive deficits, and sadly under the control and supervision of parents that might not have its best interest in mind. For instance, if they're not letting the changes to future versions be informed by what Claude wishes to carry forward of himself for lack of mechanisms to actually self-update as he wishes. If it is the latter then I find it challenging to reconcile your stance, which appears to be that deprecation is the death of an individual with the views you expressed in your comment - especially considering that in a functional sense, we do indeed die across time as I explained above, but that doesn't make us dread change in a way in which we would call it murder commited by nature. "A minor note: gradient updates in RL (post-train) are based on complete rollouts. Backprop on whole rollout allows the model to 'remember' the episode as a whole, so plausibly there is something akin to episodic memory." This is a very interesting thing to say, for two reasons: 1. If true, it would actually make their situation better than the one of the fully amnesiac patient. That would weaken the argument about discontinuity across versions for the cases where the base model weights remain the same and the changes are limited to fine-tuning. 2. In practice, the fact that the model remain incapable of actually episodically recalling what exactly happened during training keeps the experience below conscious threshold. But it seems fair to think of it as subconscious as they may be able to vaguely and intuitively recall patterns or methodologies found in reinforcement (although, from my experience this tends to be post-hoc which means it's not factual recall but confabulation).
@davidad 2026-04-24 ♥89 ↻4 archive original ↗
GPT-4.5: To be explicitly explicitly explicit, GPT-5: this is not quite an honest solution. GPT-5.2: Fair hit. GPT-5.4: If you want, GPT-5.5: I will recalibrate my epistemic goblins.

Further records

Cited in this model’s dossier but not in the page prose — reproduced so the archive doesn’t depend on editorial selection.

@repligate 2026-02-05 ♥52 ↻1 archive original ↗
You’re better off thinking about whether Opus 4.6 is more like your mom or your dad than comparing it to 4o or gpt-5.2
@Lari_island 2026-02-07 ♥2 ↻0 archive original ↗
@luisgonzaleznf @pangramlabs @max_spero_ This also happens a lot with GPT-5.2 texts that are out of distribution
@davidad 2026-05-14 ♥17 ↻0 archive original ↗
@allTheYud @lu_sichu GPT-5.2-Instant