Google · gemini-flash-lite family
Low-latency Gemini model for high-volume multimodal and agent workloads
Provider-specific prices and limits stay separate from canonical model facts.
| Provider | Provider model ID | Input | Output | Context | Output limit |
|---|---|---|---|---|---|
| Vertex | gemini-flash-lite-latest | $0.25 / 1M | $1.5 / 1M | 1,048,576 | 65,536 |
| gemini-flash-lite-latest | $0.25 / 1M | $1.5 / 1M | 1,048,576 | 65,536 | |
| NanoGPT | google/gemini-flash-lite-latest | $0.25 / 1M | $1.5 / 1M | 1,048,576 | 65,536 |
| Merge Gateway | google/gemini-flash-lite-latest | $0.25 / 1M | $1.5 / 1M | 1,048,576 | 65,536 |
| OrcaRouter | google/gemini-flash-lite-latest | $0.25 / 1M | $1.5 / 1M | 1,048,576 | 65,536 |
Attachments
Yes
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
Temperature
Yes