By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the public preview of Gemini 1.5 Pro, an advanced artificial intelligence model, on February 15, 2024. This new iteration significantly expands the model's capacity by introducing a context window of 1 million tokens. This substantial increase allows Gemini 1.5 Pro to process and analyze vastly larger amounts of information in a single prompt compared to its predecessors. For context, previous models often had context windows in the tens of thousands of tokens, limiting the depth and breadth of data they could consider simultaneously. The 1 million token context window means the model can now ingest and reason over extensive documents, hours of video, or large codebases, enabling more sophisticated analysis and understanding.
This capability is demonstrated through several examples provided by Google. The model was able to analyze a 402-page PDF document, a 44-minute silent film, and over 11 hours of audio, all within a single prompt. This showcases its ability to handle diverse and lengthy data formats. The expanded context window is powered by a new Mixture-of-Experts (MoE) architecture, which Google states makes the model more efficient. While the full details of the MoE implementation are not disclosed, it is a common technique in large language models to activate only specific parts of the neural network for a given task, thereby reducing computational overhead.
Gemini 1.5 Pro also retains the multimodal capabilities of the Gemini family, meaning it can understand and process information from various modalities, including text, images, audio, and video. The model's performance is benchmarked against previous versions, with Google reporting that Gemini 1.5 Pro demonstrates performance comparable to Gemini 1.0 Ultra, despite its more efficient architecture and larger context window. This suggests a significant leap in both capability and efficiency for Google's AI offerings. The company is making Gemini 1.5 Pro available to developers via Google AI Studio and Vertex AI, allowing them to integrate its advanced features into their applications and services.
The introduction of Gemini 1.5 Pro with its 1 million token context window positions Google at the forefront of AI development, particularly in the area of long-context understanding. This advancement has broad implications for fields requiring the analysis of extensive data, such as legal document review, scientific research, software development, and content creation. The ability to process such large volumes of information in one go could lead to more accurate insights, faster problem-solving, and the development of entirely new AI applications that were previously constrained by context limitations. Google's commitment to pushing the boundaries of AI capabilities with models like Gemini 1.5 Pro underscores the rapid pace of innovation in the artificial intelligence sector.
Original source — read the full reporting at the publisher:
Read on Fast CompanyGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.