← model directory
qwen
Qwen: Qwen3.8 Flash
qwen/qwen3.8-flash
↓ runs free on your own hardware⚙ tool calling
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
specs & pricing
- Type
- text
- Provider
- qwen
- Model ID
- qwen/qwen3.8-flash
- Capabilities
- vision, tools, reasoning
- Context window
- 1M tokens
- Self-hostable
- Yes — runs on your own GPU
- Input price
- $0.17 / 1M tokens
- Output price
- $0.52 / 1M tokens
Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.