By Interestana AI Editorial — AI-drafted, human-overseen. How we report
Google Unveils Gemini 1.5 Pro With 1 Million Token Context Window
Google announced the public preview of Gemini 1.5 Pro on February 15, 2024, featuring a groundbreaking 1 million token context window. This significant expansion allows the model to process and reason over vastly larger amounts of information than its predecessors, including entire codebases, lengthy books, or hours of video. The previous standard for large context windows in advanced models typically hovered around 128,000 to 256,000 tokens. Gemini 1.5 Pro's 1 million token capacity represents an eight-fold increase over the 128,000 token limit of Gemini 1.0 Pro, enabling it to ingest and analyze up to 1,500 pages of text, 11 hours of video, or 30,000 lines of code in a single prompt. This capability is powered by a new Mixture-of-Experts (MoE) architecture, which Google states is more efficient and performant. The MoE architecture allows the model to selectively activate specific parts of its neural network for different tasks, leading to faster processing and reduced computational cost. Google demonstrated the model's capabilities by having it analyze a 402-page PDF document and a 44-minute silent film, successfully identifying specific details and answering complex questions about their content. The company also showcased its ability to process a 1-hour YouTube video and extract specific information from it. The extended context window is particularly beneficial for developers and researchers who need to work with extensive datasets or complex documents. For instance, it can analyze entire software projects to identify bugs or suggest improvements, or process lengthy legal documents to extract key clauses. The model also maintains strong performance across various benchmarks, including multimodal reasoning tasks, despite the increased context length. Google highlighted that Gemini 1.5 Pro achieves performance comparable to Gemini 1.0 Pro on standard benchmarks, even with its significantly larger context window. The company is making Gemini 1.5 Pro available to developers via Google AI Studio and Vertex AI, with a focus on enabling them to build new applications that leverage its enhanced capabilities. The initial release is in preview, indicating that further refinements and features are expected. This advancement positions Gemini 1.5 Pro as a leading model for tasks requiring deep understanding of extensive information, potentially setting a new standard for large language model context windows in the industry.
Original source — read the full reporting at the publisher:
Read on The EconomistGet the weekly AI digest
AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.