Gemini 2.5 Flash Lite
Google🇺🇸
google/gemini-2.5-flash-liteGemini 2.5 Flash-Lite is Google's most lightweight and cost-effective model in the 2.5 family, optimized for very high-throughput, low-latency tasks such as classification, extraction, and simple chat. It trades some reasoning depth for speed and low cost.
Cost rate
0.1x
Context
1M
Released
Jul 22, 2025
Input
TextImageFileAudioVideo
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
No top-ranked categories found.
Performance
Median latency and throughput measured across recent requests.
Throughput317 tok/s
Latency