By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
Google announced the general availability of Gemini 1.5 Pro, a significant advancement in its large language model capabilities, on February 15, 2024. The most notable feature of this new model is its expanded context window, which now supports up to 1 million tokens. This substantial increase allows Gemini 1.5 Pro to process and reason over vastly larger amounts of information compared to its predecessors. For context, a 1 million token context window is equivalent to processing approximately 1,500 pages of text, or over an hour of video, or more than 30,000 lines of code. This capability is a leap forward, enabling the model to understand intricate details and long-range dependencies within extensive datasets.
Previously, models like Gemini 1.0 Pro were limited to a context window of 32,000 tokens. The jump to 1 million tokens in Gemini 1.5 Pro means users can now input and analyze much larger documents, codebases, or video files without losing crucial information. This enhanced capacity is particularly beneficial for complex tasks such as summarizing lengthy reports, analyzing extensive legal documents, debugging large software projects, or understanding the narrative arc of long videos. Google highlighted that this feature is powered by a new Mixture-of-Experts (MoE) architecture, which makes the model more efficient and scalable.
The Gemini 1.5 Pro model also retains the multimodal capabilities of its earlier versions, meaning it can process and understand various types of information, including text, images, audio, and video. The expanded context window further amplifies these multimodal abilities, allowing for deeper analysis across different data formats. For instance, users can now feed an entire hour-long video into the model and ask specific questions about its content, or provide a large codebase and request detailed explanations of its functionality or potential bugs. This level of comprehension was previously unattainable with standard context window limitations.
Google has made Gemini 1.5 Pro available through the Gemini API in Google AI Studio and Vertex AI, its cloud-based machine learning platform. The company is offering a limited preview of the 1 million token context window, with plans to expand access and potentially offer even larger context windows in the future. This development positions Gemini 1.5 Pro as a powerful tool for developers and businesses looking to leverage AI for complex analytical tasks that require processing extensive data. The introduction of this feature underscores Google's commitment to pushing the boundaries of AI model performance and utility, aiming to unlock new possibilities for information processing and understanding.
Original source — read the full reporting at the publisher:
Read on The EconomistGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.