By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
Google announced Gemini 1.5 Pro on February 15, 2024, introducing a significant advancement in artificial intelligence with its expanded context window capability. This new iteration of the Gemini model can process up to 1 million tokens, a substantial increase from the 128,000 tokens available in the previous Gemini 1.0 Pro. This expanded context window allows the AI to analyze and understand much larger volumes of information, including entire codebases, lengthy books, or hours of video content, in a single prompt.
The 1 million token context window represents a leap in the model's ability to maintain coherence and recall information across extended inputs. For developers and users, this means the AI can perform more complex tasks that require understanding intricate relationships within vast datasets. For instance, it can analyze a complete software project to identify bugs or suggest improvements, or process a full-length feature film to extract specific plot points or character arcs. Google highlighted that this capability is not just theoretical; the model has demonstrated proficiency in tasks such as summarizing lengthy documents and answering questions based on extensive technical manuals.
Gemini 1.5 Pro is built on a new Mixture-of-Experts (MoE) architecture, which Google states makes it more efficient and performant. This architecture allows the model to selectively activate different parts of its neural network for specific tasks, leading to faster processing and reduced computational overhead compared to traditional dense models. The company also emphasized that the model retains the multimodal capabilities of its predecessor, meaning it can understand and process various types of information, including text, images, audio, and video, simultaneously. This integrated approach to multimodal processing is a key feature that distinguishes Gemini from other AI models.
Initially, the 1 million token context window will be available in a limited preview for developers and enterprise customers through the Gemini API. Google plans to gradually roll out broader access and explore further applications for this advanced capability. The company also indicated that it is working on further increasing the context window size in future versions, suggesting that the current 1 million token limit is a stepping stone rather than a final destination. This development positions Google at the forefront of AI research and development, particularly in the area of large-context models, which are crucial for tackling increasingly complex real-world problems.
Original source — read the full reporting at the publisher:
Read on The EconomistGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.