← model directory

qwen

Qwen: Qwen3.8 2.4T A95B

qwen/qwen3.8-2.4t-a95b

↓ runs free on your own hardware⚙ tool calling

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

specs & pricing
Type
text
Provider
qwen
Model ID
qwen/qwen3.8-2.4t-a95b
Capabilities
tools, reasoning
Context window
1.0M tokens
Self-hostable
Yes — runs on your own GPU
Input price
$2.20 / 1M tokens
Output price
$6.60 / 1M tokens

Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.