Gemini 2.0 Flash-Lite
Gemini 2.0 Flash-Lite
📌 Model Overview
Low-latency Gemini model for high-volume multimodal and agent workloads
Gemini 2.0 Flash-Lite is a Api Only model developed by Google DeepMind, released on 2024-12-11. It is designed primarily for Multimodal General workloads. Featuring a 1.0M tokens context window and Undisclosed parameter count, it offers robust performance for enterprise integration, developers, and researchers.
✨ Key Features & Capabilities
| Feature | Description |
|---|---|
| Context Window | 1.0M tokens capacity for extended prompts and multi-turn workflows |
| Primary Task | Optimized for Multimodal General |
| Deployment | api-only |
| Modality | text, image, audio, video, pdf |
| Max Output | Max Output: 8K tokens |
| Tool / function calling support | Tool / function calling support |
⚙️ Technical Specifications
| Specification | Details |
|---|---|
| Developer / Lab | Google DeepMind |
| Release Date | 2024-12-11 |
| Model Type | Api Only |
| Parameters | Undisclosed |
| Context Window | 1.0M tokens |
| License | Proprietary |
| Model Family | gemini-flash-lite |
💰 Pricing
| Tier / Unit | Rate (USD) |
|---|---|
| 1M input tokens | $0.052 |
| 1M output tokens | $0.21 |
📜 License & Usage
This model is governed by the Proprietary license. Please check official developer guidelines before commercial deployment.
Key Features
Max Output: 8K tokens
Tool / function calling support
You might also want to compare
Verified Sources
Tags
Model Specs
Parameters
Undisclosed
Context Window
1.0M tokens
License
Proprietary
Deployment
Pricing
- 1M input tokens$0.052
- 1M output tokens$0.21
as of 2026-07-26
Lineage
Model Family
Part of the gemini-flash-lite family
Curator Notes
Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.
Compare Specs
Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.
Compare Model