OLMo 1B
OLMo 1B is worth evaluating for classification and json / tool use when its provider route and context window match the workload.
Use it for
- Teams evaluating classification and json / tool use
- Buyers comparing 1 tracked provider route
Do not use it for
- Vision or document-understanding workloads
- Family
- OLMo
- Released
- 2024-02-01
- Parameters
- 1B
- Architecture
- Decoder Only
- Knowledge cutoff
- 2023-03
- Specialization
- general
- Openness
- Open source
- License
- Apache 2.0OSI-approvedCommercial use: permitted
- Weights
- Unknown
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 1 route · OpenRouter
About
AMD OLMo 1B is a fully open-source large language model with 1 billion parameters, designed for advanced reasoning, instruction-following, and chat capabilities. Utilizing a decoder-only transformer architecture, it is trained on a 1.3 trillion-token subset of the Dolma v1.7 dataset, achieving a remarkable training throughput of 12,200 tokens per second per GPU. AMD leveraged its Instinct MI250 GPUs across 16 nodes to optimize performance, followed by a sophisticated three-stage training process of pre-training, supervised fine-tuning, and direct preference optimization.
Top use-case fit
Classification
Included by capability and metadata signals in the decision map.
JSON / Tool use
Included by capability and metadata signals in the decision map.
Provider price ladder
Compare API pricing across 1 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| OpenRouter | $15.00 | $60.00 | Serverless |
Capabilities
Benchmark peer barsfor Classification
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Cheapest of 1 route · OpenRouter