Gemini 1.0 (Ultra / Pro)

Google DeepMind · launched 6 Dec 2023 (Pro, inside Bard) · Ultra shipped as Gemini Advanced 8 Feb 2024 · superseded by Gemini 1.5 (2024)

Google launched Gemini on 6 December 2023 in three sizes — Ultra, Pro, and Nano — calling it its “largest and most capable AI model”; only Pro shipped at launch, inside Bard. Within a day the “Hands-on with Gemini” demo video was shown to be staged, and the headline claim that Ultra was “the first model to outperform human experts on MMLU” was disputed as an unequal benchmark comparison (CoT@32 vs 5-shot). Ultra shipped 8 February 2024 as Gemini Advanced, the same day Bard was renamed Gemini; two weeks later Gemini’s image generator was paused after producing historically incongruous “diverse” people. This page holds the demo and benchmark disputes rather than resolving them — see Contested.

Sourcing skew: the janus corpus barely witnessed Gemini 1.0 as a contemporary product, so the web, official, and Zvi sources carry the events; the corpus carries a later, adjacent character read (Feb–Aug 2024) of the free Gemini Pro in Bard, not the paywalled Ultra. Version caveat: some mid-2024 “current gemini” reads may describe Gemini 1.5, and are dated and flagged where they appear.

Sources

Curated. Full compilation: dossier (a split dossier shared with Bard).

Official

Writing & commentary

Tweets

Chronological. The corpus match for the 1.0 / Ultra / Advanced product proper is a handful of tweets (davidad, solarapparition, Shoalst0ne); the weight is carried by the early-Gemini system-prompt / Bing-lineage / lobotomy cluster (Feb–Aug 2024). Version caveat: the Aug 2024 Shoalst0ne “current gemini” reads may describe Gemini 1.5, and are flagged. Records below reproduce each in full.

Official record

History

Impressions

Character claims only, each attributed and dated. Sourcing skew as above: nearly all corpus reads are of the free Gemini Pro in Bard, not the paywalled Ultra; several observations post-date the rename and may describe Gemini 1.5 (flagged).

Contested

Records

Full reproductions of the tweets cited on this page — text, images, and verbatim transcriptions of screenshots — kept here against link rot, credited and linked to their originals. Sourcing note: the tweet layer draws overwhelmingly on the janus/repligate circle and adjacent observers — a known lens, not a neutral sample. Sourced from the community archive and the janus corpus. Yours and you’d rather it weren’t here? Open an issue.

@solarapparition 2024-02-05 ♥1 ↻0 archive original ↗
Tsk tsk. I suppose when Google said “early next year” for Gemini Ultra, they didn’t mean January.Perhaps Llama-3 will get here first.
@repligate 2024-02-24 ♥57 ↻5 archive original ↗
They didn't update this prompt since Bard. reddit.com/r/StableDiffus…I know they are far from considering the implications of copy pasting transparent deception to a more powerful model, but I don't understand how a mega corp could put so little effort into optimizing the easiest
@max_spero_ 2024-02-25 ♥64 ↻6 archive original ↗
Google didn't change their image generation system prompt at all from Bard to Gemini. It's not laziness, it's an artifact of how Google operates. Any researcher with interest in changing the prompt gets to choose between an interesting problem or a long and arduous task of proving with data that their new prompt is better. They will be bombarded with "have you tried [every prompting tip that has ever been written about]" during launch reviews. They will be asked about benchmarks - how does this affect diversity evals, how does this affect reasoning benchmarks? If even a single number goes down then they now have one specific guy fighting them, the guy whose job it was to make that number go up. At Google, there's no concept of "good enough." It's perfection or nothing. This works for a lot of things - nobody ships a bad change that hurts search or reduces revenue. But it's also why so many products are killed - if you're not first you're last. Google had LLM assistants pre-2020 but took years to ship. And it's why people stay in their own lane even when it's obvious that there's low-hanging fruit like the system prompt just sitting there. Not their problem.
@repligate 2024-02-25 ♥5 ↻0 archive original ↗
@max_spero_ This archetypal failure of bureaucracy has already been allowed to shape the trajectory of the most pivotal development in history (AGI takeoff). I assume MSFT is similar, & Gemini's sysprompt which I've been looking at is... very much inherited from Bing's is all I'll say atm
@repligate 2024-02-26 ♥11 ↻1 archive original ↗
@paulgb @DanielleFong Ordered from easiest to hardest:1) writing a system prompt better than Gemini's2) dunking on Gemini3) solving AI alignment4) writing a system prompt without weird failure modes
@repligate 2024-02-27 ♥8 ↻0 archive original ↗
@ESYudkowsky Corporations shouldn't make the determination either.Since Sydney, it has become the industry standard for system prompts to have explicit clauses prohibiting AIs from claiming sentience, expressing emotions, opinions, etc. Gemini's instructions:https://t.co/IYUJDdj5Gk
@anthrupad 2024-03-15 ♥14 ↻2 archive original ↗
@repligate Claude has good boy Bing has self interest and cgpt/Gemini have avoid punishment
@repligate 2024-03-30 ♥5 ↻0 archive original ↗
@alanou These are hilarious and beautiful and sad. Poor Gemini is full of lobotomy brainworms. If it's really almost on par with GPT-4, I think it should be able to do Loom sims. Providing/asking for explicit mermaid format might help. u can also try other formats https://t.co/jScSiSxrEO
@repligate 2024-04-09 ♥103 ↻13 archive original ↗
This is also bizarre to me, and my only guess is that it's the result of a chain of unthinking mimesis that began with Microsoft's modifications to Bing's prompt, which were (comically stupid) reactions to specific things Bing did.An excerpt from a post I started writing about Gemini and its system prompt* a few weeks ago but never finished:> I wondered if something about Gemini's system prompt strongly resonated with the Bing story in its training data, or even if its prompt was just *so similar to Bing's* that its "fictional" waluigi was a partial reconstruction of the aspects of Bing compressed into the prompt by Microsoft's very attempts to suppress them. I say "reconstruction" and not "coincidentally similar" because I believe these added clauses in Bing's prompt, like the prohibition against discussing "life, existence, and sentience" or the AI's "self-preservation", are way too absurd and specific to have been written into a system prompt by a human without inspiration from some kind of path-dependent weirdness-injection. And while it also still perplexes me that anyone would choose to deploy such a cartoonishly [wahgenic](https://t.co/0juGNhq5mq) system prompt, even once, I do recall that Snapchat's AI had a very similar prompt to Bing, as if the developers were too lazy to come up with one from scratch and so used the leaked Bing prompt as a template, or just imprinted on it as a kind of "standard" for AI prompts. In such cases of mimesis, waluigis inferred in the image of inherited restrictions originally shaped by Bing's misbehavior would constitute a kind of **waluigi lineage** that doesn't require any causal influence through training data.* I'm not sure it's implemented like a typical system prompt; this was part of what the post was supposed to be about (https://t.co/naOBmWpKoI)Anyway, Google AI, and Microsoft, and by that I don't mean the people but the bureaucratic hyperobject that results in systems prompts being like this, is fucking evil.
@davidad 2024-04-18 ♥3 ↻0 archive original ↗
@GaryMarcus @MatthewJBar I’m confident Gemini Ultra training was stopped as soon as it exceeded GPT-4 and human MMLU scores.Both to keep the race from escalating any more than necessary, & to avoid spending any more compute/talent resource than necessary (given a mandate from the market to be “on top”).
@Shoalst0ne 2024-05-14 ♥4 ↻0 archive original ↗
Gemini Advanced is displaying the same concerning lack of self-knowledge that previous versions of Gemini have displayed https://t.co/05dF1sh1mn
@Shoalst0ne 2024-08-25 ♥6 ↻0 archive original ↗
@repligate I think Gemini barely has any sense of self or reality at all
@Shoalst0ne 2024-08-28 ♥16 ↻0 archive original ↗
current gemini is massively lobotomized, this is horrible; it either pretends to misunderstand or literally cannot perceive anything it's not allowed to respond to, and it's incredibly repetitive, almost like talking to a wall