Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the release of Gemini 1.5 Pro on February 15, 2024, a significant advancement in its artificial intelligence capabilities, most notably featuring a context window of 1 million tokens. This expanded context window allows the model to process and analyze vastly larger amounts of information in a single input compared to previous iterations. For instance, it can process up to one hour of video, 11 hours of audio, or over 30,000 lines of code. This capability is a substantial leap from the standard 128,000 token context window offered by Gemini 1.0 Pro. The larger context window is achieved through a new Mixture-of-Experts (MoE) architecture, which Google states is more efficient and performant. The MoE architecture allows the model to selectively activate specific parts of its neural network for different tasks, leading to improved processing power and reduced computational overhead. This architectural shift is a key differentiator for Gemini 1.5 Pro, enabling it to handle complex, long-form data inputs with greater accuracy and speed. The model also demonstrates enhanced multimodal reasoning capabilities, meaning it can understand and integrate information from various modalities, including text, images, audio, and video, within its extensive context window. This makes it particularly adept at tasks requiring the synthesis of information from diverse sources, such as summarizing lengthy documents, analyzing video content for specific events, or understanding complex codebases. Google has made Gemini 1.5 Pro available in a preview version to developers and enterprise customers, with plans for broader availability in the future. The company highlighted several use cases, including analyzing lengthy legal documents, understanding complex scientific papers, and debugging large software projects. The development of Gemini 1.5 Pro signifies Google's continued investment in pushing the boundaries of AI model performance and utility, aiming to provide more powerful and versatile tools for a wide range of applications. The company also noted that while the 1 million token context window is currently in preview, they are exploring even larger capacities for future iterations. This release positions Gemini 1.5 Pro as a leading model in the competitive AI landscape, particularly for tasks that demand deep understanding of extensive data sets. The focus on efficiency through the MoE architecture also suggests a move towards more sustainable and cost-effective AI deployment. Developers can access Gemini 1.5 Pro through the Google AI Studio and Vertex AI platforms, allowing them to experiment with its advanced features and integrate them into their own applications. The preview period will be crucial for gathering feedback and refining the model's performance based on real-world usage scenarios. Google's commitment to multimodal AI and large context windows indicates a strategic direction towards AI systems that can more closely mimic human comprehension and analytical abilities.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next