Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Atlantic3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced on February 15, 2024, that its latest AI model, Gemini 1.5 Pro, now supports a context window of 1 million tokens. This significant expansion allows the model to process and analyze vastly larger amounts of information than previous iterations, including entire books, hours of video, or extensive code repositories. The previous standard for many large language models, including earlier versions of Gemini, was typically around 32,000 tokens, with some reaching up to 128,000 tokens. The 1 million token context window represents a more than eight-fold increase over the 128,000 token limit previously available in preview. This enhanced capability enables Gemini 1.5 Pro to maintain a deeper understanding and recall of information across lengthy inputs, facilitating more complex reasoning and analysis tasks. For instance, users can now input a 1,500-page document, approximately 11 hours of video, or over 30,000 lines of code and expect the model to accurately recall and reason about specific details within that content. Google highlighted this capability by demonstrating Gemini 1.5 Pro's ability to analyze a 402-page document detailing the Apollo 11 mission and answer specific questions about its content, as well as its capacity to process a 44-minute silent film and identify specific visual elements. The model also showed proficiency in analyzing a large codebase, identifying potential bugs and explaining its functionality. This advancement is built upon a new Mixture-of-Experts (MoE) architecture, which Google states makes the model more efficient and scalable. The MoE architecture allows the model to selectively activate different parts of its neural network for different tasks, leading to improved performance and reduced computational cost compared to traditional dense models. Gemini 1.5 Pro is currently available in a limited preview for developers and enterprise customers, with plans for broader availability in the future. The expanded context window is expected to unlock new applications in fields such as legal document review, scientific research, software development, and media analysis, where processing large volumes of data is crucial. Google's ongoing development of Gemini models aims to push the boundaries of AI capabilities, focusing on multimodal understanding and long-context reasoning.

Original source — read the full reporting at the publisher:

Read on The Atlantic

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next