NVIDIA

Phi 4 Multimodal — Benchmark Scores, Pricing & Performance Analysis

BUDGETNVIDIA
Input Cost
$0.00/1M
Output Cost
$0.00/1M
Context Window
128K
Max Output
16K
Accepts
text

Phi 4 Multimodal by NVIDIA demonstrates competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Cost Benchmarks

Input Cost
$0.001st of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$0.001st of 413

Cost per 1M output tokens

OpenRouter API

Context Benchmarks

Context Length
128K267th of 451

Maximum context window size

OpenRouter API

Available from 4 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 1 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Nvidia$0$0128K
NanoGPT$0.07$0.11$0.035128K
Azure$0.08$0.32128KSupported
Azure Cognitive Services$0.08$0.32128K

Phi 4 Multimodal — Frequently Asked Questions

How much does Phi 4 Multimodal cost?

Phi 4 Multimodal costs $0.00 per 1M input tokens and $0.00 per 1M output tokens. This makes it one of the more affordable models.

What is the context window of Phi 4 Multimodal?

Phi 4 Multimodal has a context window of 128K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Phi 4 Multimodal?

Phi 4 Multimodal was created by NVIDIA. It is classified as a budget model in our catalogue.

Is Phi 4 Multimodal open source?

No, Phi 4 Multimodal is a proprietary model. It is available through NVIDIA's API and compatible providers.