Gemini 1.5 Flash
Gemini 1.5 Flash is a faster, cheaper model distilled from 1.5 Pro, and 1.5 Pro's context doubled to 2M tokens (waitlist).
Flash is the fastest Gemini served in the API, trained by distillation from 1.5 Pro and keeping long context. Announced alongside Project Astra prototype agents and Gemini Nano multimodality.
- Date
- Tuesday, 14 May 2024
- Lab
- Google DeepMind
- Kind
- model
- Access
- closed API
Preview 2024-05-10 in API (changelog), I/O announcement 2024-05-14, GA 2024-05-23. 1.5 Pro 2M context reached GA 2024-06-27.
Sources
- blog.google/technology/ai/google-gemini-update-flash-ai-assistant-io-2024
- ai.google.dev/gemini-api/docs/changelog
This record was checked against its sources on 6 October 2026. How we check