By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced on February 15, 2024, the release of Gemini 1.5 Pro, an updated version of its flagship large language model, featuring a groundbreaking context window of 1 million tokens. This substantial increase from the typical 32,000 tokens found in many contemporary models allows Gemini 1.5 Pro to process and analyze vastly larger amounts of information in a single prompt. The expanded context window enables the model to understand and reason over extensive documents, lengthy codebases, and even hours of video content, marking a significant leap in AI's capacity for handling complex, multi-modal data.
This new capability means Gemini 1.5 Pro can ingest and comprehend entire books, extensive research papers, or multiple lengthy videos simultaneously. For instance, users can provide a 1,500-page book and ask specific questions about its content, or upload a two-hour movie and query its plot, characters, or specific scenes. This functionality is particularly impactful for developers and researchers who often work with large datasets or complex projects that were previously difficult for AI models to fully grasp due to context limitations.
Google highlighted the model's performance in a demonstration where Gemini 1.5 Pro successfully analyzed a 402-page PDF document, identifying a specific quote within seconds. It also processed a 44-minute silent film, accurately identifying a specific object that appeared only once. The model's ability to maintain high performance even with such extensive inputs is attributed to a new Mixture-of-Experts (MoE) architecture, which Google states makes it more efficient and scalable. This architecture allows the model to selectively activate different parts of its neural network for different tasks, optimizing processing power.
Gemini 1.5 Pro is currently available in a limited preview for developers and enterprise customers, with plans for broader availability in the future. The model will initially offer a 128,000 token context window, with the 1 million token capacity rolling out to select users. Google emphasized that this advancement is part of its ongoing commitment to pushing the boundaries of AI, making models more versatile, powerful, and useful for a wide range of applications. The development signifies a major step towards AI systems that can truly understand and interact with the complexities of human-generated data at scale.
Original source — read the full reporting at the publisher:
Read on Bon AppétitGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.