LLM Reference

Gemini 2.5 Models by Google DeepMind

Google DeepMindProprietaryHighlight
13 models2024–2025Up to 1.05m ctxFrom $0.1/1M input

Details

ResearcherGoogle DeepMind
LicenseProprietary
Commercial useCommercial use: conditional
Models13
Released2024–2025
Max context1.05m

Capabilities

Vision9 of 13 models
Multimodal9 of 13 models
Reasoning1 of 13 models
Function Calling10 of 13 models
Tool Use10 of 13 models
Structured Outputs9 of 13 models
Code Execution5 of 13 models

Links

Website

About

Gemini 2.5 is Google DeepMind's next-generation multimodal model family with enhanced reasoning, coding, and long-context capabilities.

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

11 in view2 retired

Use when the workload needs agentic, 1.05m context, and tool use.

2025-10agentic1.05m contexttool use

Use when the workload needs 1m context, tool use, and function calling.

2025-091m contexttool usefunction calling

Use when the workload needs 1m context, tool use, and function calling.

2025-071m contexttool usefunction calling

Use when the workload needs 1m context, tool use, and function calling.

2025-061m contexttool usefunction calling

Use when the workload needs 1m context, reasoning, and tool use.

2025-061m contextreasoningtool use

Use when the workload needs 1m context, structured outputs, and multimodal inputs.

2025-051m contextstructured outputsmultimodal inputs

Use when the workload needs 33k context.

2025-0433k context

Use when the workload needs audio, 128k context, and tool use.

2025-04audio128k contexttool use

Use when the workload needs audio, 128k context, and tool use.

2025-04audio128k contexttool use

Use when the workload needs audio, 128k context, and tool use.

2025-04audio128k contexttool use

Use when the workload needs 128k context, tool use, and function calling.

2024-12128k contexttool usefunction calling

Release Timeline

9 release groups
2025-12
1 retired
Gemini 2.5 Flash Live API
audio128k contexttool use
Archived
2025-10
1 current
Gemini 2.5 Pro Computer Use Preview
agentic1.05m contexttool use
Current
2025-09
1 current
Gemini 2.5 Flash Lite Preview 09-2025
1m contexttool usefunction calling
Current
2025-07
1 current
Gemini 2.5 Flash Lite
1m contexttool usefunction calling
Current
2025-06
2 current
Gemini 2.5 Flash
1m contexttool usefunction calling
Current
Gemini 2.5 Pro
1m contextreasoningtool use
Current
2025-05
1 current
Gemini 2.5 Pro Preview 05-06
1m contextstructured outputsmultimodal inputs
Current
2025-04
4 current
Gemini 2.5 Flash Live API
audio128k contexttool use
Current
Gemini 2.5 Flash TTS Preview
audio128k contexttool use
Current
Gemini 2.5 Pro TTS Preview
audio128k contexttool use
Current
2025-03
1 retired
Gemini 2.5 Pro Preview 06-05
1m contextstructured outputscode execution
Archived

Specifications(13 models)

Gemini 2.5 model specifications comparison
ModelReleasedContextVisionMultimodalReasoningFn CallingTool UseStructured OutputsCode Exec
Gemini 2.5 Pro Computer Use Preview2025-101.05mYesYesNoYesYesYesNo
Gemini 2.5 Flash Lite Preview 09-20252025-091mYesYesNoYesYesYesYes
Gemini 2.5 Flash Lite2025-071mYesYesNoYesYesYesYes
Gemini 2.5 Flash2025-061mYesYesNoYesYesYesYes
Gemini 2.5 Pro2025-061mYesYesYesYesYesYesYes
Gemini 2.5 Pro Preview 05-062025-051mYesYesNoNoNoYesNo
Nano Banana (Gemini 2.5 Flash Image)2025-0433kNoNoNoNoNoNoNo
Gemini 2.5 Flash Live API2025-04128kYesYesNoYesYesYesNo
Gemini 2.5 Flash TTS Preview2025-04128kNoNoNoYesYesNoNo
Gemini 2.5 Pro TTS Preview2025-04128kNoNoNoYesYesNoNo
Gemini Deep Research2024-12128kNoNoNoYesYesYesNo

Pricing

Gemini 2.5 model pricing by provider
ModelProviderInput / 1MOutput / 1MType
Gemini 2.5 Flash Lite Preview 09-2025GCP Vertex AI$0.1$0.4Serverless
Gemini 2.5 Flash LiteGoogle AI Studio$0.1$0.4Serverless
Gemini 2.5 Flash LiteGCP Vertex AI$0.1$0.4Serverless
Gemini 2.5 Flash Lite Preview 09-2025Google AI Studio$0.1$0.4Serverless
Gemini 2.5 Flash Lite Preview 09-2025OpenRouter$0.1$0.4Serverless
Gemini 2.5 Flash LiteOpenRouter$0.1$0.4Serverless
Gemini 2.5 Flash LiteVercel AI Gateway$0.1$0.4Serverless
Gemini 2.5 FlashGoogle AI Studio$0.3$2.5Serverless
Gemini 2.5 FlashGCP Vertex AI$0.3$2.5Serverless
Nano Banana (Gemini 2.5 Flash Image)Google AI Studio$0.3$30Serverless
Nano Banana (Gemini 2.5 Flash Image)GCP Vertex AI$0.3$30Serverless
Gemini 2.5 FlashReplicate API$0.3$2.5Serverless
Nano Banana (Gemini 2.5 Flash Image)OpenRouter$0.3$2.5Serverless
Gemini 2.5 FlashOpenRouter$0.3$2.5Serverless
Gemini 2.5 FlashVercel AI Gateway$0.3$2.5Serverless
Nano Banana (Gemini 2.5 Flash Image)Vercel AI Gateway$0.3$2.5Serverless
Gemini 2.5 Flash Live APIGCP Vertex AI$0.5Serverless
Gemini 2.5 Flash TTS PreviewGoogle AI Studio$0.5Serverless
Gemini 2.5 Pro TTS PreviewGoogle AI Studio$1Serverless
Gemini 2.5 ProGoogle AI Studio$1.25$10Serverless
Gemini 2.5 ProGCP Vertex AI$1.25$10Serverless
Gemini 2.5 Pro Computer Use PreviewGoogle AI Studio$1.25$10Serverless
Gemini 2.5 Pro Computer Use PreviewGCP Vertex AI$1.25$10Serverless
Gemini 2.5 ProOpenRouter$1.25$10Serverless
Gemini 2.5 Pro Preview 05-06OpenRouter$1.25$10Serverless
Gemini 2.5 ProVercel AI Gateway$1.25$10Serverless

Popular comparisons in this family

Frequently Asked Questions

What is Gemini 2.5 used for?
Gemini 2.5 is used for audio, agentic, and vision and multimodal work. The family description and listed model capabilities point to those workloads as the best fit.
How does Gemini 2.5 compare to T5Gemma?
Gemini 2.5 by Google DeepMind is strongest where you need audio, while T5Gemma by Google DeepMind is the closest related family to check for agent workflows and tool use. Gemini 2.5 has 13 listed variants and reaches up to 1.05m context, so compare the specs and pricing tables before choosing a production model.
Which Gemini 2.5 model should I use?
For the lowest listed input price, start with Gemini 2.5 Flash Lite through Google AI Studio at $0.1/1M input tokens. For the most capable/latest local choice, evaluate Gemini 2.5 Pro with 1m context and reasoning, tool use, function calling, structured outputs, and multimodal inputs.