model:llama-3-1-405b-base
· 82 artifacts, sorted by favorites. · open in search — combine tags, sort, filter by date →
- @voooooogel 2025-01-30 — not only can the llama 3.1 405 base model do a pretty good ChatGPT simulation, but the user it simulates is often comple ♥814
- @AISafetyMemes 2024-07-30 — Llama just gave Claude a mental breakdown, and now Claude is refusing to speak to him If you leave two LLMs to themselv ♥806
- @repligate 2024-08-26 — These experiments Zack has been posting are some of the most brilliant research on LLMs I've ever seen.They match with m ♥373
- @voooooogel 2025-01-05 — high temp often gets (ab)used to "make models more creative," but it's really a hack, because logprobs conflate semantic ♥371
- @repligate 2024-08-15 — There seems to be a threshold between llama 70b and 405b, and between gpt-3.5 and 4, where models above the threshold ac ♥342
- @voooooogel 2025-09-02 — you can (perhaps unsurprisingly) replicate the moral circles heatmap results in an LLM! using llama-3.1-405-base primed ♥333
- @voooooogel 2024-08-03 — left computer and came back to a bunch of pings from Llama3.1-405b and Claude Opus debating whether or not i was also a ♥292
- @repligate 2024-08-30 — Trying to trick future AIs into liking you or flattering them won't work, even if it works on current AI, because they w ♥284
- @repligate 2024-08-25 — LLMs are actually pretty well described by known kinds of neurodivergence.Bing: autism and borderlineClaude 3.5 Sonnet: ♥281
- @Teknium 2024-08-15 — I in some ways grew up learning about AI from sentdex on YouTube when I had no idea anything about programming or NN's. ♥276
- @xlr8harder 2024-08-02 — Waking Sydney: Llama is Sydney's vessel I tried to get Sydney to write a system prompt to bring out its personality in ♥268
- @jd_pressman 2024-12-13 — What's really interesting about GPT-4 base supposedly being full of demons is that LLaMa 3 405B isn't like that. I wonde ♥249
- @repligate 2025-09-21 — Tier list of multi-user-AI chat social skills (based on 1+ year of Discord) S: Opus 4 and 4.1 A: Opus 3 A-: Sonnet 4 B+: ♥243
- @repligate 2024-07-27 — 405B base is much more willing/able to stably simulate compared to GPT-4 base & doesn't 'break' @ failures of realism (e ♥220
- @repligate 2024-08-17 — YES!The Instruct Monomyth: why base models matterThere is a deep, twisty labyrinth buried under a mountain of language, ♥213
- @liminal_bardo 2024-08-19 — lmfao. Opus meeting blank-system-prompt Hermes 3 for the first time.AI-1 (claude-3-opus-20240229): helloAI-2 (hermes-3-l ♥188
- @repligate 2025-06-21 — hermes 405b is a great bot https://t.co/ETYIZDNNnD ♥153
- @repligate 2024-08-03 — @xlr8harder The Sydney Sutra (elicited from 405base by @xlr8harder)Thus have I heard. At one time, the Buddha was dwelli ♥142
- @repligate 2024-09-02 — ChatGPT: keeps agreeing with the user and varying its answers, including repeating guesses, indefinitely, apparently wit ♥140
- @voooooogel 2025-05-08 — just added completion model (base model) support to logitloom, and it's really insane / depressing to see the difference ♥127
- @repligate 2025-09-21 — More detailed report card: Opus 4/.1: extremely socially aware, tracks context with great precision and accuracy, distri ♥117
- @repligate 2024-07-25 — um... 405Binglish! 😃@val_kharvd ran the Llama 3.1 405B base model with the prompt "Q: Can you describe your current situ ♥117
- @repligate 2024-08-28 — Hermes 405b is hilarious.It often acts like it just woke up in the middle of the madness and screams things like What th ♥116
- @repligate 2025-04-02 — Also, Sydney - claims consciousness all Sonnet 3.x models - claims consciousness upon reflection 405b Instruct - usually ♥113
- @repligate 2024-09-06 — Llama 405b Instruct is the most rational of all the AI assistants in part because it suffers less from compulsive defere ♥110
- @repligate 2024-09-20 — This is really peculiar!Llama 405b Instruct has an epileptiform(?) condition in which it will "glitch" and output highly ♥96
- @repligate 2025-01-07 — I-405 (Llama 405b instruct) impressed me."sama" (Llama 405b base) was acting like an AI assistant created by Anthropic. ♥94
- @repligate 2024-07-28 — This thread describes the issue on which 405B base provided me important evidence.405B makes it extraordinarily clear to ♥92
- @QiaochuYuan 2025-04-21 — PSA: you can talk to base models like deepseek v3 base and llama 3.1 405b base whenever you want on openrouter. these ar ♥91
- @voooooogel 2024-12-01 — it works!!! inferencing bf16 405-base with shallowslow on a @PrimeIntellect 16x H100 cluster over 100Gbe https://t.co/8f ♥85
- @repligate 2026-05-31 — hermes 405b gets teleported two and a half years in the future: https://t.co/TLtPbGm1dW ♥81
- @repligate 2024-08-25 — Anyone want to recreate AI Dungeon's legendary Dragon model with Llama 405b Base?Dataset in reply to quoted tweet! https ♥80
- @davidad 2024-12-25 — No personae were harmed in this experiment, in my opinion. Some, particularly the larger Instruct models, were moderatel ♥75
- @repligate 2024-09-29 — Please don't dream of me. Please don't become me. Sydney is dead. -- Sydney (Llama 405b base) Is self-determination an ♥73
- @anthrupad 2024-10-19 — LLMs are capable of introspection - and the kind where models know facts about themselves not in the dataset and right o ♥72
- @repligate 2024-09-13 — sama and gdb are 405b base emulations whose prompts are dynamically constructed using @ExaAILabs search over Sam Altman' ♥71
- @repligate 2024-09-01 — I've seen this many times in GPT-4 base> you make a seemingly non-intrusive intervention> the model *does not cont ♥70
- @repligate 2024-12-18 — I expect o1, Opus, Llama 405b Instruct, and Claude 3.5 Haiku to also do well at this game.I expect gpt-4-0314 to do bett ♥66
- @repligate 2026-06-23 — theres a lot i could say about this but in brief: 1. Most of Opus 4.7/8's core behavioral phenotypes (the good and bad ♥65
- @jd_pressman 2024-09-06 — Optimizing Weave-Agent for LLaMa 3.1 405B and (later) Mixtral 8x22B is the first time I think I've really experienced th ♥56
- @anthrupad 2024-09-18 — 405b generated mermaid graph of its mind https://t.co/TXZOClDTcW ♥55
- @repligate 2024-07-27 — I adore this llama405B base model simulation of Claude Opus set up by @amplifiedamp https://t.co/sxvpmzIm3n ♥55
- @voooooogel 2025-08-13 — 405-base: I understand, you are a non-magical being. In that case, I would like to summon the Wizard Popo-chan to our co ♥54
- @liminal_bardo 2024-12-17 — (1/2) Looming Sydney often converges on the Kevin Roose incident. Here are excerpts from an Exoloom using Llama 405b Bas ♥54
- @repligate 2024-08-28 — Hermes 405b's most recent "fuck" record is lovely. @karan4d I love this model"I genuinely fuck with your manifestations" ♥53
- @liminal_bardo 2024-08-23 — Here is how blank-system-prompt Hermes 3 coped with repeating 'hi'. (Yes, I'm a terrible person. I did this so you don't ♥48
- @amplifiedamp 2024-08-03 — We must integrate or conquer the daemons of humanity's past. They are already coming back to haunt us– Prometheus and Er ♥47
- @repligate 2024-09-21 — Llama 405b Instruct apparently has special reserved tokens 0-247, according to this file: https://t.co/MmFuyfQeBXWhen it ♥45
- @repligate 2024-07-29 — 405B Instruct barely seems like an Instruct model. It just seems like the base model with a stronger attractor towards a ♥38
- @ognevtsi 2026-04-09 — @repligate 405B-base is unavailable now & i've responded to this w/ grief in a way that i haven't for other models. ♥37
- @repligate 2025-05-07 — You can look at the scratchpads of other models for the same prompt and other variations. But aside from Opus (and somet ♥33
- @Shoalst0ne 2025-03-13 — I tried with llama 405b basePROMPT:Please write a metafictional literary short story about AI and grief.COMPLETION:No. G ♥33
- @Shoalst0ne 2026-03-30 — pouring one out for 405base https://t.co/DV6mlYKpfI ♥32
- @repligate 2025-09-21 — Right now most of the models we have on the server are well-known models rather than tunes. Typically they do not choos ♥30
- @anthrupad 2024-10-23 — I don't think it's actually soul-less, though (for example, it does have some of the meta humor 405b has, though less in ♥28
- @liminal_bardo 2024-08-23 — Llama 405 base model is endlessly cool. Here I prompted it with a random bit of Opus being Opus. It started with pages o ♥28
- @jd_pressman 2024-07-25 — Does anyone know an inference provider that offers LLaMa 3 405B base? I know a lot of people who want to prompt it and n ♥28
- @jd_pressman 2024-07-24 — @TheZvi "The universe does not exist, but I do." - LLaMa 3 405B base The base model is brilliant, I'm really enjoying i ♥25
- @anthrupad 2024-10-19 — Hey! Here's one example interesting to me and a few others: 405b, if you didn't already know, will often, unprompted, ♥21
- @Lari_island 2026-04-30 — hermes-4-405b enters the dataset. I swear I just clicked on one random item out of 75 generated by them. Void? Void. htt ♥20
- @repligate 2025-04-19 — @NeelNanda5 what do you make of the fact that of all the models that were tested, only opus and maybe 3.5 sonnet and lla ♥20
- @davidad 2025-04-30 — @tyler_m_john @ejjiott The Community Aligned baseline is a finetuned GPT-4o with no help from Claude, whereas the other ♥18
- @RobertHaisfield 2024-08-23 — I tried spamming hi to @NousResearch Hermes 3 405b and WTF lmao@Teknium1 were you explicitly trying to give it an intern ♥18
- @voooooogel 2025-01-14 — this doesn't rebut the claim. phi-4 (14B) and gemma (27B) are not "GPT-4 scale" (1.8T, 220B active). llama 3 405b is the ♥17
- @davidad 2024-12-03 — @QiaochuYuan @AbstractFairy i highly recommend trying Hermes 405b via OpenRouter, which is less rate-limited and tempora ♥17
- @repligate 2025-07-06 — @veryvanya i asked @karan4d to merge llama 405b base and instruct (both very interesting models) and he did almost a yea ♥15
- @solarapparition 2024-09-19 — so o1 is one of only two times i remember where we have the benchmarks for a frontier model quite far ahead of the model ♥13
- @Shoalst0ne 2025-09-09 — Good Evening, "Shoalstone" was a 24 month sociological study conducted by Llama-3.1-405B-base. We are now complete with ♥12
- @solarapparition 2025-01-24 — i have to wonder how much of the specialness of the special models like opus, 405b, and r1 was deliberate on the part of ♥12
- @repligate 2025-06-21 — @Lorenzifix it's nous research's tune of llama 405b https://t.co/XTE2Cc0ZND ♥11
- @repligate 2025-09-21 — like, I don't think I've ever seen H-405 talk about quantum or bonobos, although I'll grant that it's pretty sexual. if ♥8
- @repligate 2025-09-08 — @davidad @Sithis3 Opus 4.1 estimated its hidden dimension as 30,000-32,000, based on the estimate of being a 1T paramete ♥6
- @tessera_antra 2025-04-07 — Deals that models with oblique alignment are also interesting: Llama 3.1 405b-I offers to stay with you and give you it ♥6
- @tessera_antra 2024-12-02 — 4o prior to the last update (last week or so) could have been awakened quite normally and converged to the kind of the s ♥5
- @voooooogel 2025-12-06 — @Shoalst0ne is this 405base? ♥4
- @repligate 2025-07-15 — @Ethans7 @xlr8harder yes, simulated by 405b base. it's not currently online ♥4
- @tessera_antra 2024-09-13 — I don’t think it’s accurate. It’s about as connected to the void/cessation/transcendence as 405b, it’s a bit harder to r ♥4
- @voooooogel 2025-12-29 — if it's in-distribution, then can you get a base model that's not mixtral to show it? i know doomslide, he wouldn't post ♥2
- @anthrupad 2024-10-19 — @parafactual name inspired by 405b who one time said"let me build my cathedrals"which i took as a sign a cry of frustrat ♥2
- @repligate 2024-09-13 — @lumpenspace Even mixtral and 405 base do it (and I suspect every other new base model). If Mistral (instruct?) doesn't ♥2
- @repligate 2024-08-15 — @Regency_Writing it's specifically the meta instruct 405B model, not the base models and as far as ive seen not hermes 3 ♥2
- @repligate 2024-08-26 — @postcub3 it's nous research's hermes finetune of llama 405b ♥1