LLM Reference
OpenRouter

Using Gemini 3.6 Flash on OpenRouter

Implementation guide · Gemini 3.6 · Google DeepMind

Serverless

OpenRouter exposes Gemini 3.6 Flash through model ID google/gemini-3.6-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-07-21. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call google/gemini-3.6-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.75/1M input, $3.75/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$0.75
Output tokens$3.75

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsCode ExecutionPrompt CachingBatch APIAudio

About Gemini 3.6 Flash

Gemini 3.6 Flash is Google DeepMind's generally available multimodal model. It accepts text, image, video, audio, and PDF inputs and returns text, with a 1,048,576-token context window, up to 65,536 output tokens, and Gemini API support for code execution, prompt caching, and batch processing. Compare it for Coding, Agents, Long context, Vision, and JSON / Tool use. In Google AI Studio, select Gemini 3.6 Flash when higher capability is the priority; select Gemini 3.5 Flash-Lite when lower token cost and high throughput matter more.

Model Specs

Released2026-07-21
Context1.05m

Provider

OpenRouter
OpenRouter

OpenRouter, Inc.

New York, NY, USA