# Gemini 1 Release

Gemini 1 was Google DeepMind's first-generation multimodal large language model family, announced on December 6, 2023, succeeding LaMDA and PaLM 2. It introduced three models - Ultra, Pro, and Nano - and powered the Gemini chatbot, marking a major step in generative AI.

Gemini 1 is a family of multimodal [large language models](https://www.wikiprompt.org/wiki/large-language-model) developed by [Google DeepMind](https://www.wikiprompt.org/wiki/google-deepmind), announced on December 6, 2023, as the successor to LaMDA and PaLM 2. It comprises three initial models: Gemini Ultra, designed for highly complex tasks; Gemini Pro, for a wide range of tasks; and Gemini Nano, for on-device tasks. The family powers the Gemini chatbot and was named after the Gemini zodiac sign, reflecting the merger of Google Brain and DeepMind. Unlike text-only predecessors, Gemini was trained to process text, images, audio, video, and computer code simultaneously, positioning it as a direct competitor to [OpenAI](https://www.wikiprompt.org/wiki/openai)'s GPT-4 and other [generative AI](https://www.wikiprompt.org/wiki/generative-ai) systems.

The release marked a pivotal moment in [artificial intelligence](https://www.wikiprompt.org/wiki/artificial-intelligence) development, as Gemini Ultra became the first model to outperform human experts on the 57-subject Massive Multitask Language Understanding (MMLU) test, scoring 90%. Google touted it as its "largest and most capable AI model," with capabilities spanning [machine learning](https://www.wikiprompt.org/wiki/machine-learning), [deep learning](https://www.wikiprompt.org/wiki/deep-learning), and [neural network](https://www.wikiprompt.org/wiki/neural-network) architectures.

## Development Background

Google first announced Gemini during the Google I/O keynote on May 10, 2023, with CEO Sundar Pichai describing it as still in early development. It was built as a collaboration between DeepMind and Google Brain, which had merged under the Google DeepMind umbrella. DeepMind CEO Demis Hassabis highlighted in interviews that Gemini would combine conversational text capabilities with AI-powered image generation, drawing on the strengths of DeepMind's AlphaGo program, which gained worldwide attention in 2016.

Development involved hundreds of engineers, with Google co-founder Sergey Brin recalled from retirement to assist, later credited as a "core contributor." The model was trained on transcripts of YouTube videos, with lawyers filtering potentially copyrighted materials. By August 2023, reports indicated a target launch of late 2023, with early access granted to select companies through Google Cloud's Vertex AI service.

## Launch and Initial Models

On December 6, 2023, Pichai and Hassabis announced "Gemini 1.0" at a virtual press conference. At launch, Gemini Pro and Nano were integrated into Bard and the Pixel 8 Pro smartphone, respectively, while Gemini Ultra was slated for "Bard Advanced" and developer availability in early 2024. The model was initially available only in English, with Google citing the need for "extensive safety testing" before wider release.

Gemini was trained on Google's Tensor Processing Units (TPUs). Gemini Ultra reportedly outperformed GPT-4, [Anthropic](https://www.wikiprompt.org/wiki/anthropic)'s Claude 2, Inflection AI's Inflection-2, Meta's LLaMA 2, and xAI's Grok 1 on industry benchmarks, while Gemini Pro outperformed GPT-3.5. Gemini Pro became available to Google Cloud customers on AI Studio and Vertex AI on December 13, 2023. In accordance with a U.S. executive order from October 2023, Google shared testing results with the federal government, and engaged with the UK government on AI safety principles from the Bletchley Park summit.

## Technical Architecture and Capabilities

Gemini's architecture leveraged [transformer](https://www.wikiprompt.org/wiki/transformer) models, incorporating [multi-head attention](https://www.wikiprompt.org/wiki/multi-head-attention) and [positional encoding](https://www.wikiprompt.org/wiki/positional-encoding) mechanisms. Its multimodal design allowed simultaneous processing of diverse data types, a departure from text-only LLMs. The model integrated techniques like [reinforcement learning from AI feedback](https://www.wikiprompt.org/wiki/rlaif) and [temperature scaling](https://www.wikiprompt.org/wiki/temperature-scaling) for output control, with [top-k sampling](https://www.wikiprompt.org/wiki/top-k-sampling) and [top-p sampling](https://www.wikiprompt.org/wiki/top-p-sampling) for generation diversity.

The model's training employed [Adam optimizer](https://www.wikiprompt.org/wiki/adam-optimizer) variants, [learning rate schedules](https://www.wikiprompt.org/wiki/learning-rate-schedule), and [gradient clipping](https://www.wikiprompt.org/wiki/gradient-clipping) to stabilize [deep learning](https://www.wikiprompt.org/wiki/deep-learning) processes. [Batch normalization](https://www.wikiprompt.org/wiki/batch-normalization) and [dropout](https://www.wikiprompt.org/wiki/dropout) were used to improve generalization, while [model pruning](https://www.wikiprompt.org/wiki/model-pruning) helped optimize efficiency for on-device deployment in Gemini Nano.

## Updates and Expansion

In January 2024, Google partnered with [Samsung](https://www.wikiprompt.org/wiki/samsung-electronics) to integrate Gemini Nano and Pro into the Galaxy S24 smartphone lineup. The following month, Bard and Duet AI were unified under the Gemini brand, with "Gemini Advanced with Ultra 1.0" released via a new "AI Premium" tier of Google One. Gemini Pro received a global launch.

In February 2024, Google launched Gemini 1.5, featuring a new architecture, a mixture-of-experts approach, and a one-million-token context window. The same month, Google debuted Gemma, a smaller, free, and open-source range of Gemini models, described as a response to Meta's open-sourcing practices. Gemini 1.5 Flash was announced on May 14, 2024, at the I/O keynote, with plans to integrate Gemini Nano into Google Chrome via its "Built-in AI" architecture. Updated models, Gemini-1.5-Pro-002 and Gemini-1.5-Flash-002, were released on September 24, 2024.

## Impact and Legacy

Gemini 1 established Google DeepMind as a leading force in [generative AI](https://www.wikiprompt.org/wiki/generative-ai), competing directly with OpenAI and [Anthropic](https://www.wikiprompt.org/wiki/anthropic). Its multimodal approach influenced subsequent model development across the industry, including efforts by [Amazon Web Services](https://www.wikiprompt.org/wiki/amazon-web-services) and [Microsoft Azure](https://www.wikiprompt.org/wiki/azure) in cloud-based AI services. The model's integration into consumer products like Pixel smartphones and Samsung devices expanded AI accessibility, while its open-source Gemma models contributed to the broader AI research community.

The release also prompted regulatory discussions, with Google engaging with U.S. and UK governments on safety testing. As of 2025, Gemini has evolved through multiple versions, with Gemini 2.0 announced in December 2024, but Gemini 1 remains significant as the foundational release that demonstrated the feasibility of large-scale multimodal AI systems.

---
Source: https://www.wikiprompt.org/wiki/gemini-1-release
License: CC BY-SA 4.0 (https://creativecommons.org/licenses/by-sa/4.0/)
Last updated: 2026-09-09T02:01:53.228165+00:00
