Qwen: Qwen3.8 2.4T A95B review: specs, pricing, and where it fits
Qwen: Qwen3.8 2.4T A95B ranks #16 on the AI Hippo board with a 1.0M context window and $2.00 input / $6.00 output per 1M tokens. Specs, cost math, and how it compares.
The quick read
Qwen: Qwen3.8 2.4T A95B from qwen is a balanced pick (#16): a 1.0M window at $2.00 input / $6.00 output per 1M tokens. It aims for capability headroom without top-tier prices—useful when neither raw cost nor absolute frontier quality is the only constraint.
Spec sheet at a glance
By the numbers: 1.0M context window; text-in / text-out; 20 exposed API parameters, including tool calling and structured outputs with reasoning support. Use these to judge fit for agentic, long-context, or structured-output workloads.
Pricing & how it compares
Pricing: $2.00 input / $6.00 output per 1M tokens. Its blended $4.00 is above the top-20 median ($0.75). On AI Hippo the composite board rewards cheap long-context models such as Meta: Llama 4 Scout, so a lower rank here means weaker value-per-token, not weaker capability.
Head-to-head
Its closest board neighbor is MoonshotAI: Kimi K3 (1.0M ctx, $9.00/1M blended). Side by side, Qwen: Qwen3.8 2.4T A95B offers 1.0M ctx at $4.00/1M blended: context is similar and, on blended token price, it is cheaper. Pick between them on whichever axis your workload is bound by.
Where it fits
Best-fit workloads: whole-repo & long-document work, tool-calling agents, structured-output pipelines, multi-step reasoning.
Cost in practice & verdict
A representative 100K-input + 20K-output task costs about $0.32 on Qwen: Qwen3.8 2.4T A95B. Verdict: a balanced mid-tier option — good when you want capability headroom without top-tier token prices.