Model Terminal

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is Google DeepMind's fastest and most cost-efficient multimodal model, optimized for high-throughput, latency-sensitive tasks such as document parsing, classification, extraction, and lightweight agentic workflows. It accepts text, image, video, audio, and PDF inputs and outputs text. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
Google AI
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed