Azure

Llama 4 Scout 17B 16E Instruct — Benchmark Scores, Pricing & Performance Analysis

OPEN SOURCE
Input Cost
$0.20/1M
Output Cost
$0.78/1M
Context Window
128K
Max Output
8K
Knowledge Cutoff
Aug 2024
Accepts
text, image

Llama 4 Scout 17B 16E Instruct by Azure demonstrates competitive pricing. View detailed benchmark data including scores across coding, math, reasoning, speed, and cost metrics.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Cost Benchmarks

Input Cost
$0.20169th of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$0.78183rd of 413

Cost per 1M output tokens

OpenRouter API

Context Benchmarks

Context Length
128K267th of 451

Maximum context window size

OpenRouter API

Available from 3 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 2 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Azure$0.2$0.78128KSupported
Azure Cognitive Services$0.2$0.78128K
Cloudflare Workers AI$0.27$0.85131KSupported

Llama 4 Scout 17B 16E Instruct — Frequently Asked Questions

How much does Llama 4 Scout 17B 16E Instruct cost?

Llama 4 Scout 17B 16E Instruct costs $0.20 per 1M input tokens and $0.78 per 1M output tokens. This makes it one of the more affordable models.

What is the context window of Llama 4 Scout 17B 16E Instruct?

Llama 4 Scout 17B 16E Instruct has a context window of 128K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Llama 4 Scout 17B 16E Instruct?

Llama 4 Scout 17B 16E Instruct was created by Azure. It is classified as a open source model in our catalogue.

Is Llama 4 Scout 17B 16E Instruct open source?

Yes, Llama 4 Scout 17B 16E Instruct is an open-source model. The model weights are publicly available for download and self-hosting.