Using Qwen3-Coder-480B-A35B-Instruct on Fireworks AI

Implementation guide · Qwen3-Coder · Alibaba

ProvisionedOpen Source

Fireworks AI exposes Qwen3-Coder-480B-A35B-Instruct through model ID accounts/fireworks/models/qwen3-coder-480b-a35b-instruct. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-06-19. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Fireworks AI and generate an API key.
  2. 2
    Use the Fireworks AI SDK or REST API to call accounts/fireworks/models/qwen3-coder-480b-a35b-instruct — see the documentation for request format.

Code Examples

Install
pip install openai
API key
FIREWORKS_API_KEY
Model ID
accounts/fireworks/models/qwen3-coder-480b-a35b-instruct

Fireworks model IDs use "accounts/fireworks/models/{model-name}" format, e.g. "accounts/fireworks/models/llama4-scout-instruct-basic" or "accounts/fireworks/models/deepseek-r1".

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["FIREWORKS_API_KEY"],
    base_url="https://api.fireworks.ai/inference/v1"
)
response = client.chat.completions.create(
    model="accounts/fireworks/models/qwen3-coder-480b-a35b-instruct",
    messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)

Pricing on Fireworks AI

Capabilities

JSON / Tool useStructured OutputsCode Execution

About Qwen3-Coder-480B-A35B-Instruct

Qwen3-Coder-480B-A35B-Instruct is Alibaba's flagship open-source code generation and agentic model, released July 22, 2025 under the Apache 2.0 license. The model has 480 billion total parameters with 35 billion active parameters per token, organized across 62 transformer layers with 160 specialized expert networks and 8 experts activated per token. It uses Grouped Query Attention (GQA) with 96 query heads and 8 key-value heads and supports a native context window of 262,144 tokens, extendable to 1 million tokens via YaRN position scaling.

Model Specs

Released2025-07-22
Parameters480B total, 35B active
Context262k
ArchitectureMixture of Experts

Provider

Fireworks AI

San Mateo, California, United States