Gemini 2.0 Flash-Lite
Low-latency Gemini model for high-volume multimodal and agent workloads.
- Capability
- —No ECI score yet
- Input
- —No listed price
- Output
- —No listed price
- Context
- 1.05M8K max output
What it costs
No per-token price is published for this model yet.
No per-token price is listed for this model on models.dev.
Independent benchmarks
Epoch AI has not published a Capabilities Index score for Gemini 2.0 Flash-Lite yet.
The details
- Lab
- Released
- Dec 11, 2024
- Knowledge cutoff
- Jun 2024
- Context window
- 1,048,576 tokens
- Max output
- 8,192 tokens
- Inputs
- Text, Images, PDFs, Audio, Video
- Output
- Text
- Reasoning
- No
- Tool calling
- Yes
- Structured output
- Yes
- Weights
- Proprietary
More from Google
About Gemini 2.0 Flash-Lite
How much does Gemini 2.0 Flash-Lite cost?
Gemini 2.0 Flash-Lite has no published per-token API price on models.dev yet.
What is the context window of Gemini 2.0 Flash-Lite?
Gemini 2.0 Flash-Lite accepts up to 1,048,576 tokens per request and can write up to 8,192 tokens in one response.
How good is Gemini 2.0 Flash-Lite?
Epoch AI has not published a Capabilities Index score for Gemini 2.0 Flash-Lite yet.
Is Gemini 2.0 Flash-Lite open source?
No. Gemini 2.0 Flash-Lite is proprietary; you use it through an API or partner platforms.
What inputs does Gemini 2.0 Flash-Lite support?
Gemini 2.0 Flash-Lite accepts text, images, PDFs, audio and video and replies in text. It is not a dedicated reasoning model, supports tool calling and can return structured JSON output.
When was Gemini 2.0 Flash-Lite released?
Google released Gemini 2.0 Flash-Lite on Dec 11, 2024. Its training data runs to Jun 2024.
What are the best alternatives to Gemini 2.0 Flash-Lite?
The closest current models from other labs on capability, price and release date are Phi-4-mini (Microsoft, $0.075 / $0.30), Nova Lite (Amazon, $0.06 / $0.24), Command R7B (Cohere, $0.037 / $0.15) and Qwen-MT Plus (Alibaba (Qwen), $2.46 / $7.37).