Alibaba

Qwen3-Omni Flash Realtime — Benchmark Scores, Pricing & Performance Analysis

MID
Input Cost
$0.52/1M
Output Cost
$2.0/1M
Context Window
66K
Max Output
16K
Knowledge Cutoff
Apr 2024
Accepts
text, image, audio, video

Qwen3-Omni Flash Realtime is an AI model by Alibaba. View detailed benchmark scores, pricing data, and performance metrics on Serenities AI Models.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

Cost Benchmarks

Input Cost
$0.52254th of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$2.0243rd of 413

Cost per 1M output tokens

OpenRouter API

Context Benchmarks

Context Length
66K352nd of 451

Maximum context window size

OpenRouter API

Available from 2 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 1 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Alibaba (China)$0.23$0.91866K
Alibaba$0.52$1.9966KSupported

Qwen3-Omni Flash Realtime — Frequently Asked Questions

How much does Qwen3-Omni Flash Realtime cost?

Qwen3-Omni Flash Realtime costs $0.52 per 1M input tokens and $2.0 per 1M output tokens. This is mid-range pricing for its capability level.

What is the context window of Qwen3-Omni Flash Realtime?

Qwen3-Omni Flash Realtime has a context window of 66K tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Qwen3-Omni Flash Realtime?

Qwen3-Omni Flash Realtime was created by Alibaba. It is classified as a mid model in our catalogue.

Is Qwen3-Omni Flash Realtime open source?

No, Qwen3-Omni Flash Realtime is a proprietary model. It is available through Alibaba's API and compatible providers.