Last refreshed 2026-09-02. Next refresh: weekly.
Why use Gemini 3.8 Flash on OpenRouter?
OpenRouter offers Gemini 3.8 Flash with pay-as-you-go pricing at $0.75/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Compare Gemini 3.8 Flash across 2 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: google/gemini-3.8-flashgoogle/gemini-3.8-flashRequest example
google/gemini-3.8-flash.Gotchas
- Use provider model ID "google/gemini-3.8-flash", not the LLMReference slug "gemini-3.8-flash".
Compare Gemini 3.8 Flash Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Google AI Studio | $0.75 | $3.75 |
| OpenRouter | $0.75 | $3.75 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.75 |
| Output tokens | $3.75 |
Capabilities
About Gemini 3.8 Flash
Gemini 3.8 Flash is Google DeepMind's generally available most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows, released September 2, 2026. It accepts text, image, video, audio, and PDF inputs and returns text, with a 1,048,576-token context window, up to 65,536 output tokens, and Gemini API support for thinking (low/medium/high; minimal is not supported), function calling, tool use, structured outputs, code execution, prompt caching, search grounding, URL context, computer use (preview), and batch, flex, and priority consumption. Official model ID: gemini-3.8-flash. Sibling of live gemini-3.7-flash under a new family gemini-3.8.