LLM Reference
OpenRouter

Using Gemini 3.7 Flash on OpenRouter

Implementation guide · Gemini 3.7 · Google DeepMind

Serverless

OpenRouter exposes Gemini 3.7 Flash through model ID google/gemini-3.7-flash. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-08-24. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call google/gemini-3.7-flash — see the documentation for request format.
  3. 3
    You'll be billed $0.75/1M input, $3.75/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$0.75
Output tokens$3.75

Capabilities

VisionMultimodalReasoningJSON / Tool useStructured OutputsCode ExecutionPrompt CachingBatch APIAudio

About Gemini 3.7 Flash

Gemini 3.7 Flash is Google DeepMind's generally available multimodal workhorse model for coding and agents, released August 13, 2026. It accepts text, image, video, audio, and PDF inputs and returns text, with a 1,048,576-token context window, up to 65,536 output tokens, and Gemini API support for thinking (low/medium/high; minimal is not supported), function calling, tool use, structured outputs, code execution, prompt caching, and batch processing. Official model ID: gemini-3.7-flash. Compare it for Coding, Agents, Long context, Vision, and JSON / Tool use.

Model Specs

Released2026-08-13
Context1.05m
Knowledge cutoff2026-03

Provider

OpenRouter
OpenRouter

OpenRouter, Inc.

New York, NY, USA