Gemini 2.0 Flash-Lite
A Gemini 2.0 Flash model optimized for cost efficiency and low latency.
Specs
- Context
- 1.05M
- Max output
- 8.2K
- Input price
- $0.075/ 1M tokens
- Output price
- $0.3/ 1M tokens
- Released
- 2025-02-25
Context
- Context window
- 1.05M
- Max input
- 1.05M
- Max output
- 8.2K
Pricing · USD / 1M tokens
- Input
- $0.075
- Output
- $0.3
- Cache read
- —
- Cache write
- —
Modalities
- Input
- TextImageVideoAudio
- Output
- Text
API
- API types
Reasoning
- Reasoning
- No
Info
- Status
- Retired
- Released
- 2025-02-25
- Knowledge cutoff
- 2024-08
Features
- Confirmed
How to call
1 provider
googlemodel = gemini-2.0-flash-lite
generate
POSThttps://generativelanguage.googleapis.com/v1beta/models/{model}:generateContent
JSON
Standard format
{
"id": "gemini-2.0-flash-lite",
"object": "model",
"created": 1740441600,
"owned_by": "google",
"name": "Gemini 2.0 Flash-Lite",
"api": {
"types": [
"generate"
]
},
"limits": {
"context": 1048576,
"input": 1048576,
"output": 8192
},
"modalities": {
"input": [
"text",
"image",
"video",
"audio"
],
"output": [
"text"
]
},
"reasoning": {
"supported": false,
"efforts": []
},
"pricing": {
"currency": "USD",
"unit": "1M_tokens",
"input": 0.075,
"output": 0.3,
"cache_read": null,
"cache_write": null
},
"features": [
"structured_output"
],
"info": {
"status": "retired",
"release_date": "2025-02-25",
"knowledge_cutoff": "2024-08",
"description": "A Gemini 2.0 Flash model optimized for cost efficiency and low latency.",
"docs": "https://ai.google.dev/gemini-api/docs/deprecations",
"verified_at": "2026-10-11"
}
}Official sources
4
Deprecations | Gemini APIai.google.dev2026-10-11存档:Gemini models | Gemini API(原网址 https://ai.google.dev/gemini-api/docs/models/gemini)web.archive.org2026-10-11存档:Gemini Developer API Pricing(原网址 https://ai.google.dev/gemini-api/docs/pricing)web.archive.org2026-10-11存档:Model versions and lifecycle | Generative AI on Vertex AI(原网址 https://cloud.google.com/vertex-ai/generative-ai/docs/learn/model-versions)web.archive.org2026-10-11