Qwen2.5-Coder-7B-Instruct on NVIDIA NIM

Qwen2.5-Coder · Alibaba

ServerlessOpen Source

Last refreshed 2026-05-19. Next refresh: weekly.

Why use Qwen2.5-Coder-7B-Instruct on NVIDIA NIM?

NVIDIA NIM offers Qwen2.5-Coder-7B-Instruct with competitive pricing. NVIDIA NIM is NVIDIA's deployment platform for GPU-accelerated inference microservices.

Compare Qwen2.5-Coder-7B-Instruct across 3 providers to find the best fit for your use case
Input / 1M
-
Output / 1M
-
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: qwen/qwen2.5-coder-7b-instruct
Model ID
qwen/qwen2.5-coder-7b-instruct

Request example

Curated snippets for this provider are not sourced yet. Use NVIDIA NIM documentation with model ID qwen/qwen2.5-coder-7b-instruct.

Gotchas

  • Use provider model ID "qwen/qwen2.5-coder-7b-instruct", not the LLMReference slug "qwen2.5-coder-7b-instruct".

Compare Qwen2.5-Coder-7B-Instruct Across Providers

ProviderInput (per 1M)Output (per 1M)
OpenRouter——
Fireworks AI$0.20$0.20
NVIDIA NIM——

Capabilities

Structured Outputs

About Qwen2.5-Coder-7B-Instruct

Instruction-optimized 7B coder matching GPT-4o on code generation and excelling in infilling and cross-file reasoning tasks.

Get Started

Model Specs

Released2024-09-19
Parameters7.61B
Context128k
ArchitectureDecoder Only
Knowledge cutoff2024-02

Related Models on NVIDIA NIM