Last refreshed 2026-05-19. Next refresh: weekly.
Why use DeciLM 7B on Microsoft Foundry?
Microsoft Foundry offers DeciLM 7B with pay-as-you-go pricing at $0.52/1M input tokens. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.
Setup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: decilm-7bdecilm-7bRequest example
decilm-7b.Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.52 |
| Output tokens | $0.67 |
Capabilities
No model capability flags are currently sourced.
About DeciLM 7B
DeciLM-7B is a cutting-edge large language model developed by Deci AI, featuring 7.04 billion parameters. This model incorporates an advanced transformer decoder architecture, utilizing variable Grouped-Query Attention (GQA) to achieve high accuracy and efficiency. The architecture is fine-tuned using Deci's proprietary Neural Architecture Search technology, AutoNAC, for optimal performance. Capable of handling sequences of up to 8192 tokens, DeciLM-7B outperforms similar or larger models across various benchmarks. An instruction-tuned version, DeciLM-7B-instruct, further enhances its capabilities, especially for instruction-following tasks. It is released under the Apache 2.0 license, making it suitable for both commercial and research use.