Using WizardLM-2 8x22B on Lepton AI API

Implementation guide · WizardLM-2 · Dreamgen

ServerlessOpen Source

Lepton AI API exposes WizardLM-2 8x22B through model ID wizardlm-2-8x22b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-29. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Lepton AI API and generate an API key.
  2. 2
    Use the Lepton AI API SDK or REST API to call wizardlm-2-8x22b — see the documentation for request format.
  3. 3
    You'll be billed $0.50/1M input, $0.50/1M output tokens. See full pricing.

Code Examples

See Lepton AI API documentation for integration details.

Pricing on Lepton AI API

TypePrice (per 1M)
Input tokens$0.50
Output tokens$0.50

Capabilities

Structured Outputs

About WizardLM-2 8x22B

WizardLM-2 8x22B, developed by WizardLM@Microsoft AI, is a powerful large language model (LLM) featuring 141 billion parameters and utilizing a Mixture of Experts (MoE) architecture. It excels in complex tasks such as chat, multilingual conversations, reasoning, and agent-based interactions. Trained with an AI-powered synthetic system incorporating techniques like Evol-Instruct and AI Align AI, the model surpasses many open-source alternatives. Despite its performance on various benchmarks, further research is essential to address potential biases and enhance reliability post "toxicity testing."

Model Specs

Released2024-01-09
Parameters8x22B
ArchitectureMixture of Experts

Provider

Lepton AI API

Lepton AI

Sacramento, California, United States