Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the public preview of Gemini 1.5 Pro on February 15, 2024, a significant advancement in large language model capabilities, most notably featuring a context window of 1 million tokens. This expanded context window allows the model to process and analyze vastly larger amounts of information in a single prompt, including entire codebases, lengthy books, or hours of video. Previously, models like Gemini 1.0 Pro had a context window of 32,000 tokens, making the 1 million token capacity a more than 30-fold increase. This enhancement is powered by Google's new Mixture-of-Experts (MoE) architecture, which enables more efficient processing of information. The MoE architecture allows the model to selectively activate specific parts of its neural network for different tasks, leading to improved performance and reduced computational cost compared to dense models of similar size. Gemini 1.5 Pro is also demonstrated to perform at a level comparable to Gemini 1.0 Ultra, Google's most capable model, across a range of benchmarks, despite being a smaller model. This suggests a more efficient and powerful architecture. The model's ability to handle such extensive context was showcased through demonstrations where it could analyze a 402-page PDF document, a 44-minute silent film, and over 11 hours of audio without losing key details. Google highlighted that this capability is particularly useful for developers and researchers who need to process and understand large datasets or complex documents. The company stated that the 1 million token context window will be available to developers via the Gemini API in Google AI Studio and Vertex AI. While the 1 million token context window is in preview, Google also mentioned that it is experimenting with a context window of up to 10 million tokens, indicating a continuous push towards even greater information processing capacity. This development positions Gemini 1.5 Pro as a leading model for tasks requiring deep understanding of extensive textual, visual, and auditory data, potentially transforming fields such as software development, legal analysis, and scientific research. The broader availability of this advanced AI model through cloud platforms signifies Google's commitment to democratizing access to cutting-edge AI technologies for businesses and developers worldwide.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next