Qwen2.5-VL-72B on Novita AI

Qwen2.5 · Alibaba

ServerlessOpen Source

Last refreshed 2026-06-29. Next refresh: weekly.

Why use Qwen2.5-VL-72B on Novita AI?

Novita AI offers Qwen2.5-VL-72B with pay-as-you-go pricing at $0.80/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.

Input / 1M
$0.80
Output / 1M
$0.80
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: qwen2.5-vl-72b-instruct
Model ID
qwen2.5-vl-72b-instruct

Request example

Curated snippets for this provider have not been sourced yet.

Gotchas

  • Use provider model ID "qwen2.5-vl-72b-instruct", not the LLMReference slug "qwen2.5-vl-72b".

Pricing

TypePrice (per 1M)
Input tokens$0.80
Output tokens$0.80

Capabilities

VisionMultimodal

About Qwen2.5-VL-72B

Qwen: Qwen2.5 VL 72B Instruct available via OpenRouter. Pricing: $0.8/1M input, $0.8/1M output.

Get Started

Model Specs

Released2025-01-26
Parameters72B
Context33k
ArchitectureDecoder Only

Related Models on Novita AI