Using DeepSeek Coder V2 Lite on Fireworks AI

Implementation guide · DeepSeek Coder V2 · DeepSeek

ServerlessProvisionedOpen Weights

Fireworks AI exposes DeepSeek Coder V2 Lite through model ID deepseek-coder-v2-lite. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-09-18. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Fireworks AI and generate an API key.
  2. 2
    Use the Fireworks AI SDK or REST API to call deepseek-coder-v2-lite — see the documentation for request format.
  3. 3
    You'll be billed $0.50/1M input, $0.50/1M output tokens. See full pricing.

Code Examples

Install
pip install openai
API key
FIREWORKS_API_KEY
Model ID
deepseek-coder-v2-lite

Fireworks model IDs use "accounts/fireworks/models/{model-name}" format, e.g. "accounts/fireworks/models/llama4-scout-instruct-basic" or "accounts/fireworks/models/deepseek-r1".

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["FIREWORKS_API_KEY"],
    base_url="https://api.fireworks.ai/inference/v1"
)
response = client.chat.completions.create(
    model="deepseek-coder-v2-lite",
    messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)

Pricing on Fireworks AI

TypePrice (per 1M)
Input tokens$0.50
Output tokens$0.50

Capabilities

No model capability flags are currently sourced.

About DeepSeek Coder V2 Lite

DeepSeek Coder V2 Lite is an open-source Mixture-of-Experts (MoE) language model specifically tailored for efficiency and cost-effectiveness in coding tasks. It operates with a 15.7B parameter count, but only 2.4B are active at any given time, making it comparable to GPT4-Turbo for code-centric applications. This model supports 338 programming languages and has an extended context length of 128K tokens, facilitating the handling of complex codebases and lengthy prompts. Its features encompass code generation, completion, understanding, and mathematical reasoning, making it versatile for diverse coding applications.

Model Specs

Released2024-06-17
Parameters16B
Context128k
ArchitectureMixture of Experts
Knowledge cutoff2023-11

Provider

Fireworks AI

San Mateo, California, United States