Interestana
Home/News/Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
The Economist3 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window

Google announced the public preview of Gemini 1.5 Pro on February 15, 2024, a significant advancement in its artificial intelligence capabilities, most notably featuring a context window of 1 million tokens. This expanded context window represents a tenfold increase over the 100,000 tokens available in the previous Gemini 1.5 Pro preview, allowing the model to process and analyze vastly larger amounts of information in a single prompt. The model can now ingest and reason over the equivalent of approximately 1,500 pages of text, or over an hour of video, or over 11 hours of audio. This capability is powered by a new Mixture-of-Experts (MoE) architecture, which Google states is more efficient and performant. The Gemini 1.5 Pro model is available to developers through the Google AI Studio and Vertex AI platforms, enabling them to build applications that leverage this enhanced context understanding. The company highlighted several use cases, including summarizing lengthy documents, analyzing codebases, and processing extensive video content for insights. For instance, developers can input entire books, lengthy research papers, or hours of meeting recordings to extract key information or identify patterns. The model's ability to handle such extensive inputs is expected to unlock new possibilities in fields like legal research, software development, and content analysis. Google emphasized that Gemini 1.5 Pro maintains the strong performance and multimodal capabilities of its predecessors, including understanding and reasoning across text, images, audio, and video. The MoE architecture, a key innovation, allows the model to selectively activate specific parts of its neural network for different tasks, leading to more efficient computation and faster processing times. This architectural shift is a departure from traditional dense transformer models and signifies Google's commitment to pushing the boundaries of AI efficiency and scalability. The company also noted that while the 1 million token context window is available in preview, they are exploring even larger capacities for future iterations. The release positions Gemini 1.5 Pro as a leading solution for tasks requiring deep comprehension of extensive data, aiming to provide developers with a powerful tool for building next-generation AI applications. The availability of this advanced model through cloud platforms underscores Google's strategy to integrate its AI innovations into enterprise solutions, facilitating wider adoption and development across various industries. The company has also committed to ongoing research and development to further enhance the model's capabilities and address potential ethical considerations associated with large-scale AI processing.

Original source — read the full reporting at the publisher:

Read on The Economist

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next