Using Mistral 7B OpenOrca on Fireworks AI
Implementation guide · OpenOrca · Alignment Lab AI
Fireworks AI exposes Mistral 7B OpenOrca through model ID accounts/fireworks/models/openorca-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.
Last refreshed 2026-04-19. Next refresh: weekly.
Quick Start
- 1
- 2Use the Fireworks AI SDK or REST API to call
accounts/fireworks/models/openorca-7b— see the documentation for request format. - 3
Code Examples
pip install openaiFIREWORKS_API_KEYaccounts/fireworks/models/openorca-7bFireworks model IDs use "accounts/fireworks/models/{model-name}" format, e.g. "accounts/fireworks/models/llama4-scout-instruct-basic" or "accounts/fireworks/models/deepseek-r1".
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["FIREWORKS_API_KEY"],
base_url="https://api.fireworks.ai/inference/v1"
)
response = client.chat.completions.create(
model="accounts/fireworks/models/openorca-7b",
messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)Pricing on Fireworks AI
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.20 |
| Output tokens | $0.20 |
Capabilities
No model capability flags are currently sourced.
About Mistral 7B OpenOrca
The OpenOrca Mistral 7B is a high-performance open-source AI model that builds upon the Mistral 7B base, known for its outstanding efficiency relative to its compact size 1. Utilizing a transformer architecture, it incorporates innovations like sliding window attention and a rolling buffer cache to reduce memory use and enhance processing on conventional GPUs 8. The model excels in tasks such as text and code generation, question answering, and dialogue, often outperforming larger counterparts on various benchmarks 2.