Using WizardLM 7B on Baseten API

Implementation guide · WizardLM · WizardLM Team

ServerlessOpen Weights

Baseten API exposes WizardLM 7B through model ID wizardlm-7b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-29. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Baseten API and generate an API key.
  2. 2
    Use the Baseten API SDK or REST API to call wizardlm-7b — see the documentation for request format.

Code Examples

See Baseten API documentation for integration details.

Pricing on Baseten API

Capabilities

No model capability flags are currently sourced.

About WizardLM 7B

WizardLM 7B is a large language model designed for effective instruction-following and has 7 billion parameters. It is developed using the Evol-Instruct method to generate a wide range of open-domain instructions and is based on the architecture of the Llama 7B model with merged delta weights for enhanced performance. The model is available in various forms, including quantized versions optimized for different hardware applications such as GGML for CPU and GPTQ for GPU inference. Some iterations are uncensored, lacking filtering for potentially harmful content, thus placing the onus on users regarding the content generated.

Model Specs

Released2023-04-28
Parameters7B
Context2k
ArchitectureDecoder Only

Provider

Baseten API

Baseten

San Francisco, California, United States