Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Details
| Field | Value |
|---|---|
| id | models/gemini-2.5-flash-lite |
| context_length | 1048576 |
| max_output_tokens | 65535 |
| quantization | unknown |