LLM Reference
OpenRouter

Using Gemini 3.8 Flash on OpenRouter

Implementation guide · Gemini 3.8 · Google DeepMind

Serverless

OpenRouter exposes Gemini 3.8 Flash through model ID google/gemini-3.8-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-09-02. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call google/gemini-3.8-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.75/1M input, $3.75/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$0.75
Output tokens$3.75

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsCode ExecutionPrompt CachingBatch APIAudio

About Gemini 3.8 Flash

Gemini 3.8 Flash is Google DeepMind's generally available most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows, released September 2, 2026. It accepts text, image, video, audio, and PDF inputs and returns text, with a 1,048,576-token context window, up to 65,536 output tokens, and Gemini API support for thinking (low/medium/high; minimal is not supported), function calling, tool use, structured outputs, code execution, prompt caching, search grounding, URL context, computer use (preview), and batch, flex, and priority consumption. Official model ID: gemini-3.8-flash. Sibling of live gemini-3.7-flash under a new family gemini-3.8.

Model Specs

Released2026-09-02
Context1.05m
Knowledge cutoff2026-03

Provider

OpenRouter
OpenRouter

OpenRouter, Inc.

New York, NY, USA