LLM Reference
Microsoft Foundry

Using Orca 2 7B on Microsoft Foundry

Implementation guide · Orca 2 · Microsoft Research

ProvisionedOpen Weights

Microsoft Foundry exposes Orca 2 7B through model ID orca-2-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-05-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Microsoft Foundry and generate an API key.
  2. 2
    Use the Microsoft Foundry SDK or REST API to call orca-2-7b — see the documentation for request format.
  3. 3
    You'll be billed $0.52/1M input, $0.67/1M output tokens. See full pricing.

Code Examples

See Microsoft Foundry documentation for integration details.

Pricing on Microsoft Foundry

TypePrice (per 1M)
Input tokens$0.52
Output tokens$0.67

Capabilities

No model capability flags are currently sourced.

About Orca 2 7B

Orca 2 7B is a large language model developed by Microsoft, focusing on reasoning tasks and providing precise single-turn responses. It is a fine-tuned version of the LLaMA-2 architecture, trained on a synthetic dataset with enhanced reasoning capabilities, moderated by Microsoft Azure content filters. While adept at handling reasoning over user-provided data, reading comprehension, math problem-solving, and text summarization, it is not optimized for chat applications without further fine-tuning. Orca 2 shows strong performance in zero-shot settings but shares some LLMs' common limitations, including biases and the potential for generating misleading content.

Model Specs

Released2023-11-21
Parameters7B
Context4k
ArchitectureDecoder Only

More Models on Microsoft Foundry

Provider

Microsoft Foundry
Microsoft Foundry

Microsoft

Redmond, Washington, United States