Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the release of Gemini 1.5 Pro, a significant advancement in its artificial intelligence model capabilities, on February 15, 2024. The core innovation of Gemini 1.5 Pro is its unprecedented 1 million token context window. This expanded context window allows the model to process and analyze vastly larger amounts of information in a single prompt compared to previous iterations. For instance, it can ingest the equivalent of over 1,500 pages of text, or approximately one hour of video, or 11 hours of audio. This capability is a substantial leap from the standard 128,000 token limit found in Gemini 1.0 Pro, and even surpasses the 200,000 token limit of some competitor models. The larger context window is achieved through a novel Mixture-of-Experts (MoE) architecture, which Google states is more efficient and performant. This architecture allows the model to selectively activate different parts of its neural network based on the input, leading to faster processing and reduced computational cost for complex tasks. Google demonstrated the model's ability to recall specific details from lengthy video files, such as identifying a particular scene in a 40-minute silent film, and to summarize extensive codebases. The company highlighted that this feature is particularly beneficial for developers working with large code repositories or researchers analyzing extensive datasets. Gemini 1.5 Pro is also being made available to developers through a private preview, with plans for a broader rollout. The model retains the multimodal capabilities of its predecessor, meaning it can understand and process information from text, images, audio, and video simultaneously. This integrated approach to multimodal understanding is a key differentiator for Google's AI development. The enhanced context window is expected to unlock new applications in areas such as long-form content analysis, complex code comprehension, and detailed video and audio processing. Google has emphasized its commitment to responsible AI development, stating that safety and ethical considerations are paramount as they expand the capabilities of their AI models. The company is implementing rigorous testing and evaluation protocols to ensure Gemini 1.5 Pro operates within safe and beneficial parameters. The introduction of Gemini 1.5 Pro positions Google at the forefront of AI innovation, particularly in the domain of large-context understanding, potentially setting a new benchmark for the industry.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next