LLM Reference

Gemini 1.5 Flash 8B

Released
2024-10-03
Last refreshed
2026-06-15
Status
Researched 105d ago
ProprietaryCommercial use: conditionalLong context

Gemini 1.5 Flash 8B is worth evaluating for long context when its provider route and context window match the workload.

Use it for

  • Teams evaluating long context
  • Workloads that can use a 1m context window
  • Buyers comparing 1 tracked provider route

Do not use it for

  • Vision or document-understanding workloads
  • Strict JSON or tool-calling flows
Specifications
Released
2024-10-03
Context
1m
Parameters
8B
Architecture
Decoder Only
Knowledge cutoff
2024-08
Specialization
general
Openness
Proprietary
License
ProprietaryCommercial use: conditional
Weights
Not released
Code
Unknown
Training
Fine-tuned
Created by

Pioneering artificial intelligence research.

London, United Kingdom
Founded 2014
Website
Pricing
Output / 1M
$0.150
Input / 1M
$0.0375

Cheapest of 1 route · GCP Vertex AI

About

Lightweight 8B variant of Gemini 1.5 Flash optimized for speed and cost-efficiency. Supports 1M token context with fast inference for real-time applications.

Top use-case fit

Long context

Included by capability and metadata signals in the decision map.

Provider price ladder

Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.

ProviderInput / 1MOutput / 1MRoute
GCP Vertex AI$0.0375$0.150
Serverless

Available via routers & gateways(13)

Capabilities

No model capability flags are currently sourced.

Benchmark peer barsfor Long context

No task-mapped benchmark peers are available for this model yet.

Migration checks

No linked migration route is available for this model yet.