Using OLMo 1B on OpenRouter

Implementation guide · OLMo · Allen Institute for Artificial Intelligence (AI2)

ServerlessOpen Source

OpenRouter exposes OLMo 1B through model ID openai/o1. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-15. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OpenRouter and generate an API key.
  2. 2
    Use the OpenRouter SDK or REST API to call openai/o1 — see the documentation for request format.
  3. 3
    You'll be billed $15.00/1M input, $60.00/1M output tokens. See full pricing.

Code Examples

See OpenRouter documentation for integration details.

Pricing on OpenRouter

TypePrice (per 1M)
Input tokens$15.00
Output tokens$60.00

Capabilities

Structured Outputs

About OLMo 1B

AMD OLMo 1B is a fully open-source large language model with 1 billion parameters, designed for advanced reasoning, instruction-following, and chat capabilities. Utilizing a decoder-only transformer architecture, it is trained on a 1.3 trillion-token subset of the Dolma v1.7 dataset, achieving a remarkable training throughput of 12,200 tokens per second per GPU. AMD leveraged its Instinct MI250 GPUs across 16 nodes to optimize performance, followed by a sophisticated three-stage training process of pre-training, supervised fine-tuning, and direct preference optimization.

Model Specs

Released2024-02-01
Parameters1B
ArchitectureDecoder Only
Knowledge cutoff2023-03

Provider

OpenRouter

OpenRouter, Inc.

New York, NY, USA