Gemini 2.0 Flash-Lite

Released
2025-02-12
Last refreshed
2026-06-29
Status
Researched 128d ago
DeprecatedProprietaryCommercial use: conditionalMultimodalRAGAgentsLong contextVisionJSON / Tool use

Gemini 2.0 Flash-Lite is a legacy integration reference; evaluate Gemini 3.1 Flash-Lite before starting new work.

Google DeepMind releases · 50 in the last 12 monthsChangelog →
Specifications
Released
2025-02-12
Context
1.05m
Architecture
Decoder Only
Specialization
general
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Training
Pretrained
Created by

Pioneering artificial intelligence research.

London, United Kingdom
Founded 2014
Website
Pricing
Output / 1M
$0.300
Input / 1M
$0.075

Cheapest of 3 routes · GCP Vertex AI

This model is deprecated. Google DeepMind recommends switching to Gemini 3.1 Flash-Lite.

About

Google Gemini 2.0 Flash-Lite is deprecated and scheduled to shut down on June 1, 2026. Use a newer low-cost Gemini model such as Gemini 3.1 Flash-Lite for current integrations. Official Gemini API pricing lists $0.075 input / $0.30 output per 1M tokens before batch discounts.

Provider price ladder

Compare all 3

Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MCacheRoute
GCP Vertex AI$0.075$0.300-
Serverless
Vercel AI Gateway$0.075$0.300read $0.020
Serverless

Available via routers & gateways(13)

Capabilities

VisionMultimodalJSON / Tool useStructured Outputs

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.