Gemini 3.1 Flash-Lite
Gemini 3.1 Flash-Lite is a released coding, rag, and agents model with 1m context; evaluate it while provider pricing coverage matures.
Use it for
- Teams evaluating coding, rag, and agents
- Workloads that can use a 1m context window
Do not use it for
- Cost-sensitive launches that need sourced token pricing
- Teams that need a tracked hosted API route today
- Family
- Gemini 3.1
- Released
- 2026-05-07
- Context
- 1m
- Architecture
- Decoder Only
- Knowledge cutoff
- 2025-01
- Specialization
- general
- Openness
- Proprietary
- License
- ProprietaryCommercial use: conditional
- Weights
- Not released
- Code
- Unknown
No tracked provider token pricing is available yet.
About
GA release of Google's most cost-efficient Gemini 3.1 model, optimized for speed, scale, and cost efficiency. Supersedes gemini-3.1-flash-lite-preview. API model ID: gemini-3.1-flash-lite. Pricing: $0.25/$1.50 per 1M tokens in/out.
Gemini 3.1 Flash-Lite is a proprietary model in the Gemini 3.1 family. The structured metadata tracks a 1m-token context window, multimodal input, function calling, tool use, structured outputs, and code execution. No headline benchmark score is tracked for Gemini 3.1 Flash-Lite yet.
Top use-case fit: coding, agents, and build tasks
Coding
Included by capability and metadata signals in the decision map.
RAG
Included by capability and metadata signals in the decision map.
Agents
Included by capability and metadata signals in the decision map.
Provider price ladder
No tracked provider token pricing is available for this model yet.
Capabilities
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Frequently asked questions
What is the context window of Gemini 3.1 Flash-Lite?
Gemini 3.1 Flash-Lite has a context window of 1m tokens.
When was Gemini 3.1 Flash-Lite released?
Gemini 3.1 Flash-Lite was released on 2026-05-07.
No tracked provider token pricing is available yet.