Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the release of Gemini 1.5 Pro, an advanced AI model, on February 15, 2024, introducing a groundbreaking 1 million token context window. This substantial increase in context length, a significant leap from previous models, allows Gemini 1.5 Pro to process and analyze vastly larger amounts of information in a single prompt. The model can now ingest and reason over entire books, lengthy codebases, or hours of video content, a capability previously limited by computational constraints and model architecture.

During a demonstration, Google showcased Gemini 1.5 Pro's ability to analyze a 402-page PDF document and a 44-minute silent film, identifying specific details and answering complex questions about their content. This functionality is attributed to a new Mixture-of-Experts (MoE) architecture, which Google states makes the model more efficient and performant. The MoE approach allows the model to selectively activate different parts of its neural network for specific tasks, optimizing resource usage and improving processing speed. This architecture is a key innovation enabling the expanded context window without a proportional increase in computational cost.

The 1 million token context window represents a paradigm shift in how AI models can interact with and understand complex, long-form data. For developers and businesses, this means the potential for more sophisticated applications in areas such as legal document review, medical record analysis, software development, and in-depth content summarization. The ability to process such extensive inputs could lead to more accurate insights and a deeper understanding of intricate datasets. Google has made Gemini 1.5 Pro available in a limited preview for developers, with plans for broader access in the coming months.

Gemini 1.5 Pro builds upon the foundation of Google's Gemini family of models, which were designed from the ground up to be multimodal, capable of understanding and operating across different types of information, including text, images, audio, and video. The introduction of the 1 million token context window specifically enhances its text and code processing capabilities, but the underlying multimodal nature of Gemini is expected to benefit from this expanded context in future iterations. The company highlighted that while the 1 million token context window is currently in preview, they are exploring even larger context windows for future development, signaling a continued commitment to pushing the boundaries of AI comprehension.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next