Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the public preview of Gemini 1.5 Pro on February 15, 2024, a significant advancement in large language model capabilities. This new model introduces a groundbreaking 1 million token context window, a substantial increase from the typical 32,000 tokens found in models like Gemini 1.0 Pro. This expanded context window allows Gemini 1.5 Pro to process and reason over vastly larger amounts of information, including entire codebases, lengthy books, or hours of video content, in a single prompt.

The 1 million token context window represents a leap forward in how AI models can understand and interact with complex data. For developers and researchers, this means the ability to feed extensive documents, code repositories, or even video files into the model for analysis, summarization, and question-answering. Google demonstrated this capability by having Gemini 1.5 Pro analyze a 402-page PDF document and a 44-minute silent film, successfully identifying specific details and answering questions about their content. The model was also shown to process 11 hours of video, 1 hour of audio, and over 30,000 lines of code within this context window.

Gemini 1.5 Pro is built on a new Mixture-of-Experts (MoE) architecture, which Google states makes it more efficient and performant than previous models. This architecture allows the model to selectively activate different parts of its neural network for specific tasks, leading to faster processing and reduced computational cost. While the 1 million token context window will be available to developers in a limited preview, Google plans to make a 128,000 token context window available more broadly in Gemini 1.5 Pro. The company also indicated that future versions could potentially support even larger context windows.

This development positions Gemini 1.5 Pro as a powerful tool for a wide range of applications, from advanced code analysis and debugging to in-depth document review and complex data interpretation. The ability to process such extensive inputs efficiently could accelerate research, software development, and content creation. Google's ongoing investment in large context windows reflects a broader trend in the AI industry towards models that can handle more information, leading to more nuanced and comprehensive understanding and output. The company is making Gemini 1.5 Pro available through Google AI Studio and Vertex AI, with plans for broader availability.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next