LLM Reference
Microsoft Foundry

Using DeciLM 7B on Microsoft Foundry

Implementation guide · DeciLM · Deci AI

ProvisionedOpen Source

Microsoft Foundry exposes DeciLM 7B through model ID decilm-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-05-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Microsoft Foundry and generate an API key.
  2. 2
    Use the Microsoft Foundry SDK or REST API to call decilm-7b — see the documentation for request format.
  3. 3
    You'll be billed $0.52/1M input, $0.67/1M output tokens. See full pricing.

Code Examples

See Microsoft Foundry documentation for integration details.

Pricing on Microsoft Foundry

TypePrice (per 1M)
Input tokens$0.52
Output tokens$0.67

Capabilities

No model capability flags are currently sourced.

About DeciLM 7B

DeciLM-7B is a cutting-edge large language model developed by Deci AI, featuring 7.04 billion parameters. This model incorporates an advanced transformer decoder architecture, utilizing variable Grouped-Query Attention (GQA) to achieve high accuracy and efficiency. The architecture is fine-tuned using Deci's proprietary Neural Architecture Search technology, AutoNAC, for optimal performance. Capable of handling sequences of up to 8192 tokens, DeciLM-7B outperforms similar or larger models across various benchmarks. An instruction-tuned version, DeciLM-7B-instruct, further enhances its capabilities, especially for instruction-following tasks. It is released under the Apache 2.0 license, making it suitable for both commercial and research use.

Model Specs

Released2024-01-16
Parameters7B
Context8k
ArchitectureDecoder Only

Provider

Microsoft Foundry
Microsoft Foundry

Microsoft

Redmond, Washington, United States