Gemini 1.5 Pro brings a 1-million-token context window
Google announced Gemini 1.5 Pro, a mixture-of-experts model with a context window of up to 1 million tokens in production preview (10M tested in research), able to process hours of video or entire codebases in a single prompt.
Key facts
- Announced 15 February 2024
- Standard 128K context; up to 1M tokens for early testers
- Research tests up to 10M tokens with near-perfect needle-in-a-haystack recall
- Mixture-of-experts architecture
- Context expanded to 2M tokens for developers in mid-2024
What happened
Google shipped a model that could reason over ~700K words, an hour of video or 11 hours of audio at once.
Why it matters
Made million-token context a practical reality and shifted how developers used LLMs (whole-document and whole-repo prompting).
Changelog
- 2026-09-29: created
Related events
Sources (2)
- officialOur next-generation model: Gemini 1.5 (Google)
- paperGemini 1.5: Unlocking multimodal understanding across millions of tokens of context (arXiv)
id: 2024-02-15-gemini-1-5 · updated 2026-09-29 · open in the interactive timeline