Microsoft Foundry exposes DeciLM 7B through model ID decilm-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-05-19. Next refresh: weekly.
Quick Start
- 1
- 2Use the Microsoft Foundry SDK or REST API to call
decilm-7b— see the documentation for request format. - 3
Code Examples
Pricing on Microsoft Foundry
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.52 |
| Output tokens | $0.67 |
Capabilities
No model capability flags are currently sourced.
About DeciLM 7B
DeciLM-7B is a cutting-edge large language model developed by Deci AI, featuring 7.04 billion parameters. This model incorporates an advanced transformer decoder architecture, utilizing variable Grouped-Query Attention (GQA) to achieve high accuracy and efficiency. The architecture is fine-tuned using Deci's proprietary Neural Architecture Search technology, AutoNAC, for optimal performance. Capable of handling sequences of up to 8192 tokens, DeciLM-7B outperforms similar or larger models across various benchmarks. An instruction-tuned version, DeciLM-7B-instruct, further enhances its capabilities, especially for instruction-following tasks. It is released under the Apache 2.0 license, making it suitable for both commercial and research use.