Model Terminal

Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite is Google DeepMind's fastest and lowest-cost model in the Gemini 2.5 family, optimized for high-throughput, latency-sensitive tasks such as classification, extraction, summarization, and routing. It supports multimodal inputs (text, image, video, audio, PDF) with a 1M-token context window and includes thinking controls, Grounding with Google Search, and code execution. Source: supabase.

Value score
Not scored
Context
1.0M
tokens
Max output
66K
tokens
Price
$0.3/1M in · $2.5/1M out

Capability radar

Identity

Developer
Google AI
Openness
Closed
Modalities
Release
2025-04-17
Knowledge cutoff
Deprecation

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
$0.3
Output / 1M
$2.5
Speed