Last refreshed 2026-05-19. Next refresh: weekly.
Why use Falcon 7B on Microsoft Foundry?
Microsoft Foundry offers Falcon 7B with pay-as-you-go pricing at $0.52/1M input tokens. Microsoft Foundry is a unified Azure platform-as-a-service offering for enterprise AI operations, model builders, and application development.
Compare Falcon 7B across 3 providers to find the best fit for your use caseSetup recipe
Docs fallbackUse the provider REST API or SDKCreate a provider API keymodel: falcon-7bfalcon-7bRequest example
falcon-7b.Gotchas
No curated gotchas have been sourced for this exact provider/model route yet.
Compare Falcon 7B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| Microsoft Foundry | $0.52 | $0.67 |
| GCP Vertex AI | — | — |
| Alibaba Cloud PAI-EAS | — | — |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.52 |
| Output tokens | $0.67 |
Capabilities
About Falcon 7B
Falcon-7B, developed by the Technology Innovation Institute, is a cutting-edge large language model boasting a decoder-only architecture with 7 billion parameters. It's trained on 1,500 billion tokens from the curated web dataset, RefinedWeb, enhancing its performance in language tasks. The model is equipped with advanced features like FlashAttention and multiquery attention, optimizing speed and memory usage. With 32 layers and rotary positional embeddings, it manages a sequence length of up to 2048 tokens efficiently.