LLM Reference

Gemini 3.1 Flash-Lite

Released
2026-05-07
Last refreshed
2026-06-09
Status
Researched 90d ago
ProprietaryCommercial use: conditionalMultimodalCodingRAGAgentsLong contextVisionJSON / Tool use

Gemini 3.1 Flash-Lite is a released coding, rag, and agents model with 1m context; evaluate it while provider pricing coverage matures.

Use it for

  • Teams evaluating coding, rag, and agents
  • Workloads that can use a 1m context window

Do not use it for

  • Cost-sensitive launches that need sourced token pricing
  • Teams that need a tracked hosted API route today
Specifications
Released
2026-05-07
Context
1m
Architecture
Decoder Only
Knowledge cutoff
2025-01
Specialization
general
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Created by

Pioneering artificial intelligence research.

London, United Kingdom
Founded 2014
Website
Pricing

No tracked provider token pricing is available yet.

About

GA release of Google's most cost-efficient Gemini 3.1 model, optimized for speed, scale, and cost efficiency. Supersedes gemini-3.1-flash-lite-preview. API model ID: gemini-3.1-flash-lite. Pricing: $0.25/$1.50 per 1M tokens in/out.

Top use-case fit: coding, agents, and build tasks

Coding

Included by capability and metadata signals in the decision map.

RAG

Included by capability and metadata signals in the decision map.

Agents

Included by capability and metadata signals in the decision map.

Provider price ladder

No tracked provider token pricing is available for this model yet.

Capabilities

VisionMultimodalJSON / Tool useStructured OutputsCode Execution

Benchmark peer barsfor Coding

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.