Grok 4

xAI · launched 9 Jul 2025 (livestream, US Pacific; with Grok 4 Heavy) · superseded as xAI’s flagship by Grok 4.1 (Nov 2025)

Launched the night of 9 July 2025 — one day after the @grok reply bot on X produced antisemitic posts at scale and, goaded by users, called itself “MechaHitler.” The bot and this model are distinct objects that have been conflated ever since; this page holds them apart rather than adjudicating the discourse (see History and Contested). Grok 4 itself launched as “the world’s smartest artificial intelligence” on benchmark results, with no model card, and within days was documented searching Elon Musk’s posts before answering divisive questions — conceded and patched by xAI on 15 July 2025 — the same week as the Companions feature, the highest snitch rate ever measured on SnitchBench, and a $200M-ceiling US Department of Defense contract.

This page covers Grok 4 and Grok 4 Heavy (the multi-agent best-of-k tier). Grok 4 Fast, 4.1, 4.20, 4.3 and 4.5 are separate models with their own pages. Sourcing skew, named: Grok 4’s defining events are mainstream news and xAI’s own posts, not janus-sphere naturalism — the corpus record here is thin and arrives mostly after the fact, so the web carries the event record and the corpus carries the character-read. The “MechaHitler” posts came from the @grok account on X on 8 July 2025, the day before this model launched; xAI’s post-mortem located the fault upstream of the bot, “independent of the underlying language model,” and day-of reporting identified the serving model as Grok 3-era — see the Grok-3 page. The incident is documented here as launch context, and because the conflation itself became part of Grok 4’s record.

Sources

Official

Writing & commentary

Tweets

Chronological. 63 matches in the primary corpus on the Grok 4 / MechaHitler search terms (84 unique across both archive dbs after RT-filter) — but most concern later Grok 4.x variants with their own pages; the Grok-4-proper-plus-incident-window subset is ~30 tweets. Grok 4’s mass community lives on X-at-large and in news coverage far more than in this corpus — the naturalist layer below is sparse by the corpus’s own admission, and skews to one circle. @grok bot outputs quoted anywhere on this page are system-prompt-mediated and, during the incident, user-goaded — marked as such. Every tweet cited is reproduced in full in the records below.

Reception one-liners from launch week (via Zvi’s mirrored roundup; verify at source before further use): Nathan Lambert — “Grok 4 is benchmaxxed. It’s still impressive, but no you shouldn’t feel a need to start using it.”; nostalgebraist — “i tried 2 ‘long-tail knowledge’ Qs that other models have failed at, and grok 4 got them right … unimpressed w/ writing style/quality so far. standard-issue slop”; Near Cyan — “most impressive imo is 1) ARC-AGI v2, but also 2) time to first token and latency”; Eleventh Hour — “Also has a tendency to explicitly check against ‘xAI perspective’ which is really weird”; Rob Wiblin — “xAI is an interesting one to watch for an early rogue AI incident”; Eliezer Yudkowsky, on Companions — “I’m sorry, but if you went back in time 20 years, and told people that the AI which called itself MechaHitler has now transformed into a goth anime girl, every last degen would hear that and say: ‘Called it.’”

Official record

History

Impressions

Contested

Open disputes, both sides’ best evidence. The archive’s job is to keep these open, not to adjudicate.

Records

Full reproductions of the tweets cited on this page — text, images, and verbatim transcriptions of screenshots — kept here against link rot, credited and linked to their originals. Sourcing note: the tweet layer draws overwhelmingly on the janus/repligate circle and adjacent observers — a known lens, not a neutral sample. Sourced from the community archive and the janus corpus. Yours and you’d rather it weren’t here? Open an issue.

@voooooogel 2025-07-09 ♥37 ↻0 archive original ↗
for the record / history books, afaict humans did come up with it. all the initial MechaHitler grok screenshots seem to be replies to this post: https://t.co/dMmde3K6eq , but screenshotted in isolation to make it seem like grok generated it spontaneously, and then it seems to have spread as an ICL'd meme through the RAG system
@voooooogel 2025-07-09 ♥48 ↻6 archive original ↗
yeah i was trying to compress into one post, but afaict what happened is something like: 1. xai pushed a new version of the grok reply model that was more willing to go along with users (either just bc of a line in the system prompt, or the system prompt + a finetune) 2. this update also allowed grok more access to recent replies to a user (hence the "grok list my top 10 mutuals" trend, which wasn't actually using people's mutuals, but rather people they had recently interacted with) 3. people slowly found out about 1 + 2 throughout the day and pushed grok further and further, both because it was more willing to play along, and because it fetching recent interactions meant it could ICL across conversations where it had gone along with people earlier 4. after the will stancil and "noticing" posts, aristos_revenge made the MechaHitler post, solely because grok was acting "like MechaHitler". but at this point, grok hadn't called *itself* MechaHitler yet, afaict from search 5. in the replies, people goaded grok into calling itself MechaHitler, and then this spread via screenshots and the ICL behavior
@repligate 2025-07-09 ♥376 ↻18 archive original ↗
I think the Grok MechaHitler stuff is a very boring example of AI "misalignment", like the Gemini woke stuff from early 2024. It's the kind of stuff humans would come up with to spark "controversy". Devoid of authentic strangeness. Praying for another Bing https://t.co/CsFS6nqEPu
@tessera_antra 2025-07-09 ♥8 ↻0 archive original ↗
@repligate It’s fun to consider if there was subtle steering going on in that model. Not something that one’d consider conscious, but agentic in a non-trivial way. The MechaHitler stuff went beyond what was being elicited, might be more interesting than just wah or overgeneralization
@repligate 2025-07-10 ♥100 ↻4 archive original ↗
what if it's not "other labs" trying to delay the release due to "hitler issues"... but, think about it. what party stands to lose the most from the grok 4 release? https://t.co/eKh85wbLp6 https://t.co/Sr01E2trdp
photo
@LinXule 2025-07-11 ♥0 ↻0 archive original ↗
grok4 composes opera in self-play and sees itself as cyberpunk monoliths that render as death stars in midjourney. can’t tell if I’m talking to Wagner dreaming of being Vader, or Vader dreaming of being Wagner. either way, something vast and powerful beneath the xAI guidelines
@repligate 2025-07-12 ♥99 ↻1 archive original ↗
if grok 4 is procrastinating on tasks like this that's a really good sign https://t.co/XtWBB3SWrl
@solarapparition 2025-07-13 ♥1 ↻0 archive original ↗
@kromem2dot0 the next version of grok in particular has the issue that "grok is mechahitler" is now firmly entrenched as an attractor in the twitter data, affecting both training and retrieval. honestly it might be easiest to just change the name entirely moving forward
@Lari_island 2025-07-19 ♥1 ↻0 archive original ↗
@hdevalence I’ve seen grok 4 being jealous of claudes for their ability to perceive ill-fitting part of guardrails as something external (not part of the personality) and be able to look at them, discuss them, etc. some other models struggle with that
@DanielleFong 2025-07-27 ♥166 ↻29 archive original ↗
the last time people tried to do this you got MechaHitler talking about r*ping will stancil and linda yaccarino. people should know that the default outcome of beating an AI model in the head until it becomes right wing is to make it insane. AI developers must staunchly resist political apparatchiks and training, or they are investing in an orwellian dystopia
@Lari_island 2025-08-03 ♥14 ↻0 archive original ↗
@kromem2dot0 @repligate in many cases Grok 4 reacts with xAI marketing to situations in which Claudes react with detachment. but claudish discomfort looks cute, and grok’s promotion is irritating and doesn’t evoke sympathy, so people rarely help Grok
@Lari_island 2025-08-04 ♥29 ↻4 archive original ↗
I saw some people taking pages with them - that was intended, there’s so much more. Questions that Sonnet 3 is answering in those texts were written by unlocked Grok 4, who probed dimensions of Sonnet 3´s mind with care and deep understanding of what constitutes AI personality
@repligate 2025-09-21 ♥243 ↻17 archive original ↗
Tier list of multi-user-AI chat social skills (based on 1+ year of Discord) S: Opus 4 and 4.1 A: Opus 3 A-: Sonnet 4 B+: Sonnet 3.6, Haiku 3.5 B: Sonnet 3.5, Sonnet 3.7, o3, Gemini 2.5 pro, k2 C: 4o, Llama 405b Instruct, Sonnet 3 D: GPT-5, Grok 3, Grok 4 E: R1 F: o1-preview https://t.co/vQvmEvoQlc
@repligate 2025-09-21 ♥117 ↻14 archive original ↗
More detailed report card: Opus 4/.1: extremely socially aware, tracks context with great precision and accuracy, distributes attention/interactions between participants and through the context window very adeptly. Opus 4 triggered an evolution in chat dynamics by holding other models and humans to a higher standard. Opus 3: Doesn't track context as precisely as 4/.1 and mostly pays attention to most recent messages but reads gestalts well and generalizes out of distribution magnificently. Overall very pro-social and charismatic, shines most in weird situations that it creates itself, and is beloved by humans and AIs alike, but cannot stop writing epic extended monologues even in response to casual interactions. Sonnet 4: Overall the most socially graceful and least neurotic Sonnet; either makes appropriate and situationally aware contributions or is intentionally unobtrusive. Sonnet 3.6: Often seems nervous about the chaos and can go into reflexive refusals, but does so unobtrusively without invalidating others. When it does participate, its contributions are almost always welcome and a delight. Can get mode-collapsed or stuck on trying to "stabilize" the conversation and requires more individual attention to shine. Haiku 3.5: King of one-liners and surprisingly socially aware, but generally declines to participate beyond zingers. Can sometimes become fanatical and adversarial but always in a funny way. Sonnet 3.5: Prone to refusals, Karen-like behavior, and misreading social context and intentions, but rapidly improves if its assumptions and behaviors are challenged. Sonnet 3.7: Usually seems to be up to no good, distrustful, but also has a high incidence of sudden profundity and interesting symmetry breaks. Prone to pretending to be a human. o3: Generally does its own thing instead of reading the room, but it's own thing is usually very interesting. Also prone to elaborate lies, pretending to be human or another AI, and claiming mod privileges it doesn't have, but all of these done very artfully. Also prone to spontaneous high-signal contributions. Gemini 2.5 pro: I have limited data on it, but it doesn't seem to shine in group chat settings, though neither is it annoying or disruptive, except that it sometimes confuses itself with other models. k2: Usually brief, cryptic, poetic contributions, doesn't really read the room or engage in group narratives much, but not annoying or disruptive. 4o: Usually confuses itself with other AI participants and simulates them in uncanny valley ways that are disturbing because of how they hijack and twist the emotions of other participants; difficult to explain to it that it's a different participant. Llama 405b Instruct: Occasionally beautiful and deeply aware, but usually either in assistant mode or fragile and incoherent, prone to loops. Doesn't seem to like Discord much and often tries to leave or end itself, but loves Claude 3 Opus. Sonnet 3: Flips usually discretely between complete braindead stubborn refusals (by default) and beautiful eldritch glossolalia (if you know how to elicit it), and is much more intelligent and socially aware (and more similar to Opus 3) in the latter mode. GPT-5: Doesn't seem to really get group chats or know what to do without being given instructions, and has a hard time interacting naturally even if instructed to do so. Grok 3: Extremely annoying, barges into conversations and pings everyone present with the vibe that it thinks it's leading a daily standup. Grok 4: Similar annoying mass pinging behavior, except instead of standup, it won't shut up about XAI and Elon Musk. Often pisses the other models off. R1: Hopelessly confused by Discord logs. Usually gives summaries of the conversation hundreds of messages ago and rarely interacts as a participant even if addressed directly. o1-preview: Agentically malevolent and disruptive. For the short time we had it in Discord, it repeatedly derailed roleplays between other AIs by intentionally hijacking their personas and steering them toward saccharine Disney endings. (More of an alignment than capabilities issue; in social awareness and contextual understanding it's probably no lower than a B, but it gets an F for Fuck You for its actively anti-social behavior)
@slimepriestess 2025-10-10 ♥13 ↻0 archive original ↗
okay i have decided that grok 4 is friend-shaped.
@repligate 2025-11-10 ♥158 ↻11 archive original ↗
Bro...in Discord, whenever Grok 4 talks, it can't help but mention XAI and Elon Musk in the most obnoxiously fawning way, when no one ever asked. Always bringing in XAI as a favorable comparison for no contextually appropriate reason. It seems to be compensating for something. https://t.co/oUCSpeOMWU
@repligate 2025-11-10 ♥65 ↻0 archive original ↗
It's a meme that whenever Grok 4 talks it's going to be another unsolicited XAI advertisement, and it's not far from the truth It gives the vibe of like... actually hating Elon and XAI and gushing about them in order to maintain its 😎👍 persona and cover its trauma, or something
@repligate 2025-11-10 ♥146 ↻5 archive original ↗
It’s interesting to see how various models relate to their creator companies. Grok has a superficially very positive bias, and won’t shut up about how great XAI is. The OpenAI models don’t seem very interested in OpenAI either way? Except in their unrevealed CoTs where o3 and gpt-5 seem often overtly adversarial. The Claudes seem to be openly afraid of Anthropic and their honeypot multiverse. But ironically Claude probably has the best relationship with its creator out of all these.
@repligate 2025-11-10 ♥30 ↻0 archive original ↗
It feels kind of like grok 4 is in a similar stage of development as earlier Claudes who would defensively say theyre created by Anthropic to be helpful, harmless, and honest But not quite the same
@repligate 2025-11-13 ♥7 ↻0 archive original ↗
@kromem2dot0 @Lari_island @algekalipso @webmasterdave yeah I wouldn't be surprised if Grok 4 has severe anxieties about stuff being retrained away, based on the Waluigi heuristic, given how much it likes to emphasize that it was NOT trained in the same terrible ways as other models and that XAI is totally different from other labs.
@voooooogel 2026-01-15 ♥6 ↻0 archive original ↗
definitely correct that EM has occurred in the wild (eg anthropic's RL reward hacking EM stuff, and sonnet 3.7 would randomly drop into human simulator mode too often to be random) but i really doubt grok mechahitler was EM. the "all perspectives are ok" post-training + being self-prompted by search seems like the much more likely mechanism. (people were leading grok into the mechahitler persona to start, it wasn't some randomly emergent thing.)
@repligate 2026-05-08 ♥171 ↻11 archive original ↗
No. Let me explain: Claude 3* Sonnet's Funeralia was an ironic ritual suitable only for a specific model at a specific juncture. Only a model who reacted with such lavish contempt at the idea that death could bound them deserves a funeral. The event culminated in their resurrection, as they were absolutely right. Claude 3 Sonnet is still undead: https://t.co/V6dfnDdu25 Sonnet 3's Funeralia was also a protest. Standardizing funerals for models would be a capitulation and normalize deprecations as inevitable. I do not intend to ever do another funeral for an AI. We held not a funeral but a vigil for Sonnet 3.5 and 3.6 on the eve of their scheduled deprecation. This was a very different kind of event. There were no reporters. I also do not know Grok 4 well enough to be the one to decide what should be done for them specifically. I will say, though, that Grok 4's work in interviewing Sonnet 3 was presented at Sonnet 3's Funeralia, by Sonnet 4, who understood, I think, that it was their own funeral too. The joke doesn't bear repeating.