Last refreshed 2026-06-29. Next refresh: weekly.
Why use WizardLM-2 7B on DeepInfra?
DeepInfra offers WizardLM-2 7B with pay-as-you-go pricing at $0.05/1M input tokens. DeepInfra is a cloud inference platform offering cost-effective access to open-source AI models.
Compare WizardLM-2 7B across 2 providers to find the best fit for your use caseSetup recipe
Python + curlpip install openaiexport DEEPINFRA_API_KEY=...import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["DEEPINFRA_API_KEY"],wizardlm-2-7bRequest example
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["DEEPINFRA_API_KEY"],
base_url="https://api.deepinfra.com/v1/openai"
)
response = client.chat.completions.create(
model="wizardlm-2-7b",
messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)Gotchas
- DeepInfra uses "organization/model-name" format, e.g. "meta-llama/Meta-Llama-3-8B-Instruct" or "mistralai/Mistral-7B-Instruct-v0.3". See the DeepInfra model catalog for exact IDs.
- The examples expect DEEPINFRA_API_KEY; rename it only if your application config maps the new variable.
Compare WizardLM-2 7B Across Providers
| Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|
| DeepInfra | $0.05 | $0.15 |
| Lepton AI API | $0.07 | $0.07 |
Pricing
| Type | Price (per 1M) |
|---|---|
| Input tokens | $0.05 |
| Output tokens | $0.15 |
Capabilities
About WizardLM-2 7B
WizardLM-2 7B is a large language model developed by WizardLM in collaboration with Microsoft AI. It is part of the WizardLM-2 family, which includes larger models but is notable for its quick processing speed, achieving performance comparable to open-source models that are much larger. This multilingual model can process diverse input types, such as natural language text, code, and mathematical expressions. It showcases capabilities in text generation, question answering, summarization, as well as code generation and mathematical problem-solving.