NVIDIA

Qwen2.5 Coder 32b Instruct — Benchmark Scores, Pricing & Performance Analysis

OPEN SOURCENVIDIA
Input Cost
$0.66/1M
Output Cost
$1.0/1M
Context Window
33K
Max Output
4K
Accepts
text

Qwen2.5 Coder 32b Instruct is an AI model by NVIDIA. View detailed benchmark scores, pricing data, and performance metrics on Serenities AI Models.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Cost Benchmarks

Input Cost
$0.66267th of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$1.0197th of 413

Cost per 1M output tokens

OpenRouter API

Context Benchmarks

Context Length
33K369th of 451

Maximum context window size

OpenRouter API

Available from 7 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 3 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Nvidia$0$0128K
Hugging Face$0.06$0.2131KSupported
Alibaba (China)$0.287$0.861131K
Cloudflare Workers AI$0.66$133KSupported
Eden AI$0.66$133K
Kilo Gateway$0.66$133K
OpenRouter$0.66$133KSupported

Qwen2.5 Coder 32b Instruct — Frequently Asked Questions

How much does Qwen2.5 Coder 32b Instruct cost?

Qwen2.5 Coder 32b Instruct costs $0.66 per 1M input tokens and $1.0 per 1M output tokens. This is mid-range pricing for its capability level.

What is the context window of Qwen2.5 Coder 32b Instruct?

Qwen2.5 Coder 32b Instruct has a context window of 33K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Qwen2.5 Coder 32b Instruct?

Qwen2.5 Coder 32b Instruct was created by NVIDIA. It is classified as a open source model in our catalogue.

Is Qwen2.5 Coder 32b Instruct open source?

Yes, Qwen2.5 Coder 32b Instruct is an open-source model. The model weights are publicly available for download and self-hosting.