By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
Google announced the public preview of Gemini 1.5 Pro on February 15, 2024, featuring a groundbreaking 1 million token context window. This significant expansion allows the AI model to process and analyze vastly larger amounts of information, including entire codebases, lengthy books, and hours of video content, within a single prompt.
The enhanced context window represents a 10x increase over the previous 100,000 token limit of Gemini 1.0 Pro. This capability is powered by Google's new Mixture-of-Experts (MoE) architecture, which the company states is more efficient and performant. The MoE approach allows the model to selectively activate relevant parts of its network for specific tasks, leading to faster processing and reduced computational cost.
During its preview, Gemini 1.5 Pro demonstrated its ability to analyze a 402-page document in under 2 seconds and a 44-minute silent film in 38 seconds. This performance highlights the practical applications for developers and researchers, enabling them to gain insights from extensive datasets that were previously unmanageable. The model's multimodal capabilities, including understanding video and audio, are also enhanced by this larger context window.
Google emphasized that while the 1 million token context window is available in the preview, a standard tier offering 128,000 tokens will be available for general use. The company is also exploring further increases to the context window size in the future. This development positions Gemini 1.5 Pro as a leading model for complex reasoning and information retrieval tasks, setting a new benchmark in the field of large language models.
Original source — read the full reporting at the publisher:
Read on The EconomistGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.