By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
Google announced the release of Gemini 1.5 Pro on February 15, 2024, a significant advancement in its artificial intelligence model capabilities. The most notable feature of Gemini 1.5 Pro is its expanded context window, which can now accommodate up to 1 million tokens. This represents a substantial increase from the standard context windows of most contemporary AI models, which typically range from 32,000 to 128,000 tokens. A token is a unit of text or code that an AI model processes, and a larger context window allows the model to consider and retain more information from a given input. This enhanced capacity means Gemini 1.5 Pro can process and reason over much larger documents, extensive codebases, or hours of video content without losing track of earlier details. For instance, it can analyze an entire novel, a lengthy research paper, or a full-length movie in a single pass. This capability is crucial for complex tasks such as summarizing lengthy reports, answering detailed questions about extensive datasets, or understanding nuanced narratives within video content. The model also incorporates a Mixture-of-Experts (MoE) architecture, which Google states makes it more efficient and faster to train and run compared to traditional dense models of similar size. This architecture allows the model to selectively activate different parts of its neural network for specific tasks, optimizing performance and resource utilization. Gemini 1.5 Pro is built on the same architecture as the original Gemini models, which were designed from the ground up to be multimodal, capable of understanding and operating across different types of information including text, images, audio, and video. The 1 million token context window is available in a preview version for developers, who can access it through the Google AI Studio and Vertex AI platforms. Google has indicated that this feature will be rolled out more broadly over time. The development of Gemini 1.5 Pro signifies a push towards more powerful and versatile AI systems that can handle increasingly complex and data-intensive real-world applications. This leap in context window size is expected to unlock new possibilities in areas like scientific research, legal document analysis, software development, and entertainment content creation, where the ability to process vast amounts of information is paramount. The company also highlighted that the model maintains a high level of performance and accuracy despite the expanded context, a testament to the underlying architectural improvements. The availability of this advanced AI model to developers is a key step in fostering innovation and the development of next-generation AI-powered applications.
Original source — read the full reporting at the publisher:
Read on The EconomistGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.