Last refreshed 2026-07-21. Next refresh: weekly.
Why use Gemini 3.5 Flash-Lite on OpenRouter?
OpenRouter offers Gemini 3.5 Flash-Lite with pay-as-you-go pricing at $0.30/1M input tokens. OpenRouter is a multi-provider LLM aggregator offering unified API access to 300+ models from all major labs and emerging providers, with automatic failover for reliability.
Compare Gemini 3.5 Flash-Lite across 2 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: google/gemini-3.5-flash-litegoogle/gemini-3.5-flash-liteRequest example
google/gemini-3.5-flash-lite.Gotchas
- Use provider model ID "google/gemini-3.5-flash-lite", not the LLMReference slug "gemini-3.5-flash-lite".
Compare Gemini 3.5 Flash-Lite Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Google AI Studio | $0.30 | $2.50 |
| OpenRouter | $0.30 | $2.50 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.30 |
| Output tokens | $2.50 |
Capabilities
About Gemini 3.5 Flash-Lite
Gemini 3.5 Flash-Lite is Google DeepMind's generally available multimodal model in the Gemini 3.5 family. It accepts text, image, video, audio, and PDF inputs and returns text, with a 1,048,576-token context window, up to 65,536 output tokens, and Gemini API support for code execution, prompt caching, and batch processing. Compare it for Coding, Agents, Long context, Vision, and JSON / Tool use. In Google AI Studio, select Gemini 3.5 Flash-Lite when lower token cost and high throughput are the priority; select Gemini 3.6 Flash when higher capability matters more.