Skip to content
Sovyron

Qwen3-Omni Flash Realtime

qwen66k context16k max outputreleased Sep 15, 2025Tool calling

As of Oct 4, 2026, the cheapest Qwen3-Omni Flash Realtime API price is $0.23 per 1M input tokens (Alibaba (China)) and $0.918 per 1M output tokens (Alibaba (China)). Qwen3-Omni Flash Realtime is offered by 2 providers in the qwen family. Alibaba's own rate is $1.99 per 1M output tokens.

Cheapest input
$0.23
Cheapest output
$0.918
Official (Alibaba)
$1.99
output per 1M tokens
Providers
2
offering this model
Spread
2.2×
max / min output

API pricing by provider

Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge. How prices are ranked.

Qwen3-Omni Flash Realtime API prices by provider in USD per 1M tokens, updated Oct 4, 2026
Provider1M input1M output1M cache read1M cache writeContext
Alibaba (China)Cheapest$0.23Cheapest$0.918——66k
Alibaba$0.52$1.99——66k

Qwen3-Omni Flash Realtime pricing FAQ

What is the cheapest Qwen3-Omni Flash Realtime API?

As of Oct 4, 2026, Alibaba (China) has the lowest Qwen3-Omni Flash Realtime output price at $0.918 per 1M tokens, and Alibaba (China) has the lowest input price at $0.23 per 1M tokens. Prices are metered per-token rates in USD; free tiers and subscription plans are excluded.

How much does Qwen3-Omni Flash Realtime cost on Alibaba?

Alibaba charges $0.52 per 1M input tokens and $1.99 per 1M output tokens for Qwen3-Omni Flash Realtime.

How many providers offer Qwen3-Omni Flash Realtime?

2 providers list Qwen3-Omni Flash Realtime on Sovyron; 2 of them sell it at a metered per-token price.

Is Qwen3-Omni Flash Realtime free?

No provider in the Sovyron catalog lists a free tier for Qwen3-Omni Flash Realtime.

What is the context window of Qwen3-Omni Flash Realtime?

Qwen3-Omni Flash Realtime supports a 66k-token context window and up to 16k output tokens.