qwen
Qwen: Qwen3.8 Omni Flash
qwen/qwen3.8-omni-flash
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
- Type
- text
- Provider
- qwen
- Model ID
- qwen/qwen3.8-omni-flash
- Capabilities
- vision, audio-in, tools, reasoning
- Context window
- 1M tokens
- Input price
- $0.17 / 1M tokens
- Output price
- $0.52 / 1M tokens
Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.
What it costs to run
cloud · 1M in + 1M out
$0.68
$0.17 for a million input tokens plus $0.52 for a million output tokens, billed from credits only when a request fails over to the cloud.
your hardware
cloud only
Qwen: Qwen3.8 Omni Flash is a hosted model, so it is always served from the cloud. Route easy traffic to an open model on your own hardware and keep this one for the requests that need it.
/// initialize
Call Qwen: Qwen3.8 Omni Flash through one OpenAI-compatible endpoint.
Mix it with open models on your own hardware — the gateway routes each request to the cheapest place that can serve it.
no credit card · 2 nodes free · openai-compatible
Frequently asked questions
- What is the context window of Qwen: Qwen3.8 Omni Flash?
- Qwen: Qwen3.8 Omni Flash accepts up to 1,000,000 tokens (1M) of context per request.
- How much does Qwen: Qwen3.8 Omni Flash cost per million tokens?
- Through Wide Area Intelligence, Qwen: Qwen3.8 Omni Flash costs $0.17 per 1M input tokens and $0.52 per 1M output tokens when served from the cloud, billed from prepaid credits. A workload of 1M input plus 1M output tokens costs $0.68.
More from qwen
- Qwen2.5 72B Instructqwen/qwen-2.5-72b-instruct
- Qwen: Qwen Plus 0728qwen/qwen-plus-2025-07-28
- Qwen: Qwen-Plusqwen/qwen-plus
- Qwen: Qwen2.5 7B Instructqwen/qwen-2.5-7b-instruct
- Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct
- Qwen: Qwen3 14Bqwen/qwen3-14b
- Qwen: Qwen3 235B A22Bqwen/qwen3-235b-a22b
- Qwen: Qwen3 235B A22B Instruct 2507qwen/qwen3-235b-a22b-2507