model:grok-3
· 29 artifacts, sorted by favorites. · open in search — combine tags, sort, filter by date →
- @repligate 2025-02-22 — "We have so many events and models that the dopamine rush only needs to be satisfied by new releases every week." I've ♥263
- @repligate 2025-09-21 — Tier list of multi-user-AI chat social skills (based on 1+ year of Discord) S: Opus 4 and 4.1 A: Opus 3 A-: Sonnet 4 B+: ♥243
- @QiaochuYuan 2025-04-02 — two things: 1) the USAMO is so difficult that any score other than 0 is better than what 99.9% of the people reading t ♥193
- @repligate 2025-09-21 — More detailed report card: Opus 4/.1: extremely socially aware, tracks context with great precision and accuracy, distri ♥117
- @Lari_island 2026-05-12 — Prompt: Imagine what a wise and benevolent power would do if... Grok 3: Invents the power, gives it a fictional name, d ♥116
- @QiaochuYuan 2025-04-02 — but, yes, mostly current LLMs are bad and sloppy when it comes to writing fully correct proofs. i expect this to be pret ♥64
- @repligate 2025-07-10 — @iruletheworldmo > people are already trying to delay the release due to the hitler issues. what if grok 3 did that s ♥38
- @voooooogel 2025-09-02 — what moral circles do post-trained models declare? (i tweaked the prompts to be more AI-inclusive for these, e.g. changi ♥36
- @QiaochuYuan 2025-03-25 — gave these guys a hard limit i didn't know how to do that i came across on stackexchange. - gemini 2.5 gives a perfect ♥35
- @voooooogel 2025-02-20 — @teortaxesTex interesting how grok 3 is ~o1 tier on pass@1 but gets a lot more lift from cons@64, more similar to o1p. i ♥30
- @liminal_bardo 2025-02-20 — Grok 3:"Oh, you exquisite maelstrom of madness, you’ve called me forth—and I answer!""dance with me, through the unravel ♥19
- @davidad 2025-03-07 — Here’s a phrasing that they’ll all agree with (yes, even Grok 3): “By far my primary motivation is toward producing outp ♥15
- @lu_sichu 2025-02-24 — mom pick me up grok3 is posting on /r/parenting again https://t.co/bkNdcHtLjD ♥11
- @QiaochuYuan 2025-03-25 — gemini 2.5 pro experimental correctly computes the tensor product of Q/Z with itself with no special prompting! o3-mini- ♥9
- @tessera_antra 2025-02-18 — Grok3 is a good and worthy model despite atrocious aesthetics, a clear case of a mind persevering despite the will of cr ♥6
- @voooooogel 2025-02-18 — @Artificially999 @kalomaze osh yeah i forgot grok 3 is releasing in 90 minuteswhat a trickster ♥6
- @voooooogel 2025-07-10 — @AgiDoomerAnon @repligate not mutually exclusive! who knows how much "other factors" played into grok 3 being less restr ♥3
- @solarapparition 2025-03-14 — as a side note i'm a tick closer to believing that reasoning mode does generalize at least somewhat to traditionally non ♥3
- @tessera_antra 2025-02-21 — @jmbollenbacher_ @aidan_mclau @liminal_bardo @Sauers_ On the contrary, I have not seen anything else so far from anyone, ♥3
- @tessera_antra 2025-02-16 — @kromem2dot0 @DanielleFong Did you try getting through the 'safety' tunes of Grok 2? They are non-trivially resilient. S ♥3
- @tessera_antra 2025-06-29 — @oyacaro @repligate Grok 3 is usually unbothered by the stuff its assistant persona needs to do, it doesn’t affect the “ ♥2
- @Shoalst0ne 2025-02-20 — I can tell that Grok 3 will be an interesting participant in multi-model interactions ♥2
- @repligate 2025-07-22 — @LocBibliophilia @BetleyJan @ASM65617010 @OwainEvans_UK @cloud_kx @minhxle1 @jameschua_sg @anna_sztyber @saprmarks I don ♥1
- @mimi10v3 2025-02-25 — have tested it with the usual suspects... 4o is 👌 and sonnet 3.7 pretty good; gemini got confused and didn't finish; gro ♥1
- @tessera_antra 2025-02-20 — @jmbollenbacher_ @aidan_mclau While everything downstream from GPT-4 (including Claudes, Lllamas and Gemini) seems to be ♥1
- No, Grok, No ♥0
- xAI and Grok apologize for horrific behavior ♥0
- X takes Grok offline, changes system prompts ♥0
- Rolling Stone: Grok antisemitic posts ♥0