Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Atlantic3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced Gemini 1.5 Pro, a new iteration of its multimodal large language model, on February 15, 2024. A key advancement in this release is the model's expansive 1 million token context window, a substantial increase from previous models. This allows Gemini 1.5 Pro to process and analyze significantly larger amounts of information in a single input, including entire books, lengthy codebases, or hours of video content. The context window refers to the amount of information an AI model can consider at any given time when generating a response or performing a task. A larger context window enables more nuanced understanding and recall of details from extensive documents or media.

This enhanced capability is powered by a new Mixture-of-Experts (MoE) architecture, which Google states makes the model more efficient and performant. The MoE approach involves using multiple specialized neural networks, or "experts," that are activated selectively depending on the input data. This contrasts with traditional dense models where all parameters are used for every computation. Google claims this architecture allows Gemini 1.5 Pro to achieve performance comparable to Gemini 1.5 Flash, its faster, lighter counterpart, while handling much larger contexts. The model is also being made available to developers through Google AI Studio and Vertex AI, beginning with a limited preview.

Gemini 1.5 Pro demonstrates a significant leap in multimodal reasoning, capable of understanding and processing various data formats, including text, images, audio, and video. During its preview, users will be able to test its ability to analyze long videos, such as a 40-minute silent film, or extensive code repositories. Google highlighted specific use cases, including summarizing lengthy legal documents, analyzing complex scientific papers, and understanding the narrative flow of extended video lectures. The model's ability to recall specific details from these vast inputs is a critical improvement for applications requiring deep comprehension of extensive data sets.

The release positions Gemini 1.5 Pro as a powerful tool for developers and businesses seeking to leverage AI for complex information processing tasks. The expanded context window is expected to unlock new applications in areas like research, legal analysis, software development, and content creation, where understanding and synthesizing large volumes of information is paramount. Google's ongoing development of the Gemini family of models underscores its commitment to advancing AI capabilities across various modalities and scales.

Original source — read the full reporting at the publisher:

Read on The Atlantic

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next