Fast multimodal model with 1M context
Gemini 2.5 Flash delivers near-Pro performance at a fraction of the cost. Its 1M-token context window and low latency make it the go-to model for real-time applications, summarisation, and document processing at scale.
Try Gemini 2.5 FlashSourced from official model cards and academic papers. Higher is better.