Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
by Google|1M context|$0.05/M input tokens|$0.20/M output tokens
Endpoints
Available providers for this model, with details on pricing, context limits, and real-time health metrics.