AI Research Atlas

Gemini 1.5 Flash

Google DeepMind · 14 May 2024

Gemini 1.5 Flash is a faster, cheaper model distilled from 1.5 Pro, and 1.5 Pro's context doubled to 2M tokens (waitlist).

Flash is the fastest Gemini served in the API, trained by distillation from 1.5 Pro and keeping long context. Announced alongside Project Astra prototype agents and Gemini Nano multimodality.

Date
Tuesday, 14 May 2024
Lab
Google DeepMind
Kind
model
Access
closed API

Preview 2024-05-10 in API (changelog), I/O announcement 2024-05-14, GA 2024-05-23. 1.5 Pro 2M context reached GA 2024-06-27.

Sources

  1. blog.google/technology/ai/google-gemini-update-flash-ai-assistant-io-2024
  2. ai.google.dev/gemini-api/docs/changelog

This record was checked against its sources on 6 October 2026. How we check

Related