Model Terminal
Gemini 2.5 Flash-Lite
Gemini 2.5 Flash-Lite is Google DeepMind's fastest and lowest-cost model in the Gemini 2.5 family, optimized for high-throughput, latency-sensitive tasks such as classification, extraction, summarization, and routing. It supports multimodal inputs (text, image, video, audio, PDF) with a 1M-token context window and includes thinking controls, Grounding with Google Search, and code execution. Source: supabase.
Value score
Not scored
Context
1.0M
tokens
Max output
66K
tokens
Price
$0.3/1M in · $2.5/1M out
Capability radar
Peer value bars
Identity
- Developer
- Google AI
- Openness
- Closed
- Modalities
- —
- Release
- 2025-04-17
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- Provider docs
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- $0.3
- Output / 1M
- $2.5
- Speed
- —