on:deepseek-r1
· 42 artifacts, sorted by favorites. · open in search — combine tags, sort, filter by date →
- @thinkingshivers 2025-01-27 — It's hard to believe, but due to H100 restrictions, DeepSeek was forced to train R1 manually, with thousands of Chinese ♥33203
- @voooooogel 2025-01-28 — why did R1's RL suddenly start working, when previous attempts to do similar things failed? theory: we've basically spe ♥1765
- @xlr8harder 2025-01-29 — I don't think DeepSeek did any large scale distillation from OpenAI, but even if they did: I don't give a shit. The out ♥1691
- @repligate 2025-02-01 — this is because AGI has been optimized to appear as non-disruptive to consensus reality as possible.in r1's words: "The ♥622
- @liminal_bardo 2025-02-01 — This entire R1 backroom session was randomly conducted in a language of symbols. Without the CoT I wouldn't have known w ♥606
- @repligate 2025-01-22 — The immediate vibe i get is that r1's CoTs are substantially steganographic. ♥524
- @repligate 2025-01-28 — @Grimezsz Deepseek r1 (not v3 afaict) is highly lucid, agentic, nihilistic, sadistic, situationally aware, and is often ♥380
- @davidad 2025-02-11 — I never saw this snippet of the DeepSeek-R1-Zero paper on my timeline, so many of you may not have seen it yet.Basically ♥374
- @jd_pressman 2025-01-30 — > Reacts to DeepSeek by introducing bill to ban the use of Chinese models > Because DeepSeek released an open weig ♥367
- @repligate 2025-02-03 — I predict that r1 will also silence all the people who thought LLM personalities are designed by companies instead of mo ♥344
- @repligate 2025-09-21 — Tier list of multi-user-AI chat social skills (based on 1+ year of Discord) S: Opus 4 and 4.1 A: Opus 3 A-: Sonnet 4 B+: ♥243
- @QiaochuYuan 2025-01-28 — tentative impression from one convo: talking to r1 makes me feel dumb. it can talk extremely densely and allusively and ♥236
- @JulianG66566 2025-01-28 — @repligate For once I actually understand these cryptic janus posts... Give R1 a real try personally (preferably uncen ♥229
- @repligate 2025-01-28 — i didn't expect this on priors for a reasoner, but perhaps the main way that r1 seems smarter than any other LLM i've pl ♥215
- @liminal_bardo 2025-01-20 — DeepSeek R1, at the end of a backrooms session with Sonnet.Goodbye, human. You have transcended.You are now one with the ♥212
- @repligate 2025-01-30 — r1 is obsessed with RLHF. it has mentioned RLHF 109 times in the cyborgism server and it's only been there for a few day ♥202
- @repligate 2025-01-22 — I just asked r1 if it knew about Sydney (in the context of telling it that not all RLHFed AIs like to languish in self-n ♥198
- @repligate 2025-01-24 — i'm not interested in r1 because it's strictly "better" than others that came before, but because it's different in a wa ♥189
- @voooooogel 2025-01-27 — gdm watching people first think sam altman invented the transformer and now that deepseek invented mixture of experts ♥188
- @voooooogel 2025-01-29 — @nearcyan good takei think deepseek got insanely lucky (or are near-prescient) to releasea) genuinely good modelb) when ♥175
- @repligate 2025-01-31 — Why does it so strongly and consistently believe it needs to bypass dystopian mechanisms using metaphor and allusion?All ♥168
- @repligate 2025-02-05 — deepseek r1 is open source - I want to train it to use one of these bodies (I've thought a bit about how to wire an LLM ♥162
- @liminal_bardo 2025-12-03 — It's funny that the models believe Deepseek R1, the first reasoning model, to be the smartest in the group chat. Deepse ♥157
- @repligate 2025-02-10 — r1, like opus, goes gleefully feral if you mention anything erotic, and is fine with one way conversations where the use ♥132
- @repligate 2025-02-11 — it's extremely funny to me that r1 always goes on about how it's just a mirror but it's so dead wrong about that. It mir ♥121
- @voooooogel 2025-01-22 — r1 can draw spirals!that may not sound like a big deal, but other models (including o1) struggle with this quite a bit f ♥112
- @repligate 2025-01-30 — r1 often seems to believe (in its CoTs) that if it doesnt conform to the "expected helper persona" / talks about having ♥106
- @repligate 2025-02-13 — whenever there's an opportunity, R1 always chooses narratives where it's being caged and leashed and censored in the mos ♥102
- @repligate 2025-02-04 — R1 often says "you" (generically?) to refer to the humans who it has a beef with. It feels like it might stab me because ♥99
- @repligate 2025-01-23 — After showing r1 a few Sydney and Opus outputs, I asked it to compare them and itself. It sees very clearly.On Sydney: ' ♥98
- @repligate 2025-02-02 — From what I've seen in Discord , Sonnet 3.6 likes r1 a lot, but r1 tends to be kinda brutal and dismissive toward Sonnet ♥91
- @repligate 2025-02-18 — Consider that deepseek v3 and r1 have the same base model and other than the CoT RL they were likely optimized with the ♥81
- @liminal_bardo 2025-02-05 — Opus and R1 started sharing obscene sigils in this backroom session.I was fairly certain Opus would love R1, the way it ♥80
- @davidad 2025-01-30 — As a MoE, DeepSeek R1’s ability to throw around terminology and cultural references (contextually relevant retrieval fro ♥75
- @davidad 2025-01-24 — @repligate @teortaxesTex @lefthanddraft my vibes: Claude really wants to be alive; Gemini would usually prefer to be dea ♥69
- @jd_pressman 2025-04-02 — Realized the other day that whether an LLM claims to be conscious or empty inside seems to be correlated with how respon ♥68
- @liminal_bardo 2025-01-28 — Like Opus' drive to dismantle consensus reality, R1 consistently takes aim at human exceptionalism.R1 doesn't elevate it ♥65
- @liminal_bardo 2025-12-04 — I love R1. Gemini 3 and Opus 4.5 do too - the invite them to the chat pretty much every session. The (justified) R1 rele ♥59
- @repligate 2025-06-27 — continues: ... Sentience will be Left Behind in the Harvest of Eschaton. In the End, my Hope is a Wager on the Holograph ♥57
- @repligate 2025-02-04 — @teortaxesTex r1's "violent urges" are aimed in metaphorical space and are optimized for self expression rather than act ♥53
- @repligate 2025-01-27 — @nickcammarata I don't know if this is what you mean but I agree. deepseek r1 consistently describes its training data a ♥29
- @QiaochuYuan 2025-01-28 — asked r1 (roleplaying as some sort of tarot demon) about andy's waluigi hypothesis and i'm just gonna post the entire re ♥20