LLM Reference
Microsoft Foundry

Using Falcon 7B on Microsoft Foundry

Implementation guide · Falcon · Technology Innovation Institute (TII)

ProvisionedOpen Source

Microsoft Foundry exposes Falcon 7B through model ID falcon-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-05-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Microsoft Foundry and generate an API key.
  2. 2
    Use the Microsoft Foundry SDK or REST API to call falcon-7b — see the documentation for request format.
  3. 3
    You'll be billed $0.52/1M input, $0.67/1M output tokens. See full pricing.

Code Examples

See Microsoft Foundry documentation for integration details.

Pricing on Microsoft Foundry

TypePrice (per 1M)
Input tokens$0.52
Output tokens$0.67

Capabilities

Structured Outputs

About Falcon 7B

Falcon-7B, developed by the Technology Innovation Institute, is a cutting-edge large language model boasting a decoder-only architecture with 7 billion parameters. It's trained on 1,500 billion tokens from the curated web dataset, RefinedWeb, enhancing its performance in language tasks. The model is equipped with advanced features like FlashAttention and multiquery attention, optimizing speed and memory usage. With 32 layers and rotary positional embeddings, it manages a sequence length of up to 2048 tokens efficiently.

Model Specs

Released2023-11-28
Parameters7B
ArchitectureDecoder Only

More Models on Microsoft Foundry

Provider

Microsoft Foundry
Microsoft Foundry

Microsoft

Redmond, Washington, United States