LLM Reference

Gemini 1.5 Flash on Google Vertex AI (Extended Context)

Released
2024-02-15
Last refreshed
2026-06-15
Status
Researched 120d ago
ProprietaryCommercial use: conditionalMultimodalRAGLong contextVisionJSON / Tool use

Gemini 1.5 Flash on Google Vertex AI (Extended Context) is worth evaluating for rag, long context, and vision when its provider route and context window match the workload.

Use it for

  • Teams evaluating rag, long context, and vision
  • Workloads that can use a 1m context window
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Workloads where another current model has stronger sourced task evidence
Specifications
Released
2024-02-15
Context
1m
Architecture
Decoder Only
Knowledge cutoff
2024-05
Specialization
general
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Created by

Pioneering artificial intelligence research.

London, United Kingdom
Founded 2014
Website
Pricing
Output / 1M
$0.210
Input / 1M
$0.070

Cheapest of 1 route · GCP Vertex AI

About

Gemini 1.5 Flash on Google Vertex AI (Extended Context) is Google DeepMind's Gemini 1.5 model with multimodal text and image input. It offers a 1M-token context window.

Top use-case fit

RAG

Included by capability and metadata signals in the decision map.

Long context

Included by capability and metadata signals in the decision map.

Vision

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
GCP Vertex AI$0.070$0.210
Serverless

Available via routers & gateways(13)

Capabilities

VisionMultimodalStructured Outputs

Benchmark peer barsfor RAG

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.