Qwen3.8 2.4T A95B pricing in 2026
What one task costs on Qwen3.8 2.4T A95B
Per-million-token prices are hard to feel, so here are typical token counts for common jobs priced with Qwen3.8 2.4T A95B. Figures are USD at the rates above and exclude prompt caching and batch discounts.
| Task | Tokens (in / out) | Cost |
|---|---|---|
| Code review of a 500-line pull request | 60,000 / 4,000 | $0.144 |
| Debug one failing test | 40,000 / 6,000 | $0.116 |
| Generate a landing page in one shot | 3,000 / 30,000 | $0.186 |
| Draft a 1,200-word article | 2,000 / 2,500 | $0.019 |
| Answer a question from a 20-page document | 25,000 / 800 | $0.055 |
| Multi-step agent run (10 tool calls) | 250,000 / 25,000 | $0.650 |
| All of the above, once | $1.17 |
Cheaper at the same or better capability: GLM-5.3 Flash runs the same set of tasks for $0.091 instead of $1.17.
Qwen3.8 2.4T A95B by Alibaba costs $2 per 1M input tokens and $6 per 1M output tokens — about 200.0k tokens per $1. Prices last verified 2026-09-19.
The verdict
Alibaba's Qwen3.8 open-weight frontier (index 40) at $2 / $6. A capable open-weight model with strong front-end design, but its subscription usage burns extremely fast — practitioners report eating through 57% of a plan in just a few prompts.
A concrete example: sending 100,000 input tokens and getting back 20,000 output tokens with Qwen3.8 2.4T A95B costs about $0.32 at current prices.
Who is Qwen3.8 2.4T A95B for?
Qwen3.8 (the 2.4T-A95B open-weight release) is Alibaba's capable open-weight flagship — an index of 40 at $2 / $6 per 1M with strong front-end design. The proprietary Qwen3.8 Max sits alongside it at the same price.
Its main problem is usage burn: practitioners report eating through subscription allowances extremely fast, and it's not a value leader on per-token price. For open-weight design work it's excellent; for budget, DeepSeek Flash or GLM deliver more per dollar.
Strengths & weaknesses
- Strength: Frontier open-weight design
- Best for: Design, Open-weight
- Weakness: Subscription usage burns fast; not a value leader on per-token price.
Qwen3.8 2.4T A95B key facts
| Provider | Qwen pricing |
|---|---|
| Input price | $2 / 1M tokens |
| Output price | $6 / 1M tokens |
| Tokens / $1 | 200.0k |
| Intelligence / $1 | 80,000 |
| Context window | 1.0M tokens |
| Tier | frontier |
| Speed | Slow |
| Intelligence index | 40 / 100 |
| Released | 2026-08-12 |
| Last verified | 2026-09-19 |
What Qwen3.8 2.4T A95B delivers for your budget
For a coding workload, here is roughly what each monthly budget buys:
| Budget | Tokens / month | Months of workload |
|---|---|---|
| $20 | 6.8M | 10.5x |
| $50 | 17.1M | 26.3x |
| $100 | 34.2M | 52.6x |
Other Qwen models
How Qwen3.8 2.4T A95B compares to its closest rivals
Rather than repeating the full catalog on every page, here are the three current models nearest to Qwen3.8 2.4T A95B on the intelligence index — and why you might still pick one instead:
- DeepSeek V4.1 Flash (DeepSeek) — $0.15 input / $0.60 output per 1M tokens. Why you might pick it: Best intelligence-per-dollar on the site.
- Gemini 3.7 Flash (Gemini) — $0.75 input / $3.75 output per 1M tokens. Why you might pick it: Frontier-adjacent index at Flash price.
- Gemini 3.8 Flash (Gemini) — $0.75 input / $3.75 output per 1M tokens. Why you might pick it: Near-Pro quality at Flash price and speed.
Qwen3.8 2.4T A95B — frequently asked questions
How much does Qwen3.8 2.4T A95B cost?
Qwen3.8 2.4T A95B by Alibaba costs $2 per 1M input tokens and $6 per 1M output tokens.
How many tokens per dollar does Qwen3.8 2.4T A95B give you?
On a typical 1:3 input-to-output mix, Qwen3.8 2.4T A95B delivers around 200.0k tokens per dollar — an intelligence-per-dollar score of 80000.
What is Qwen3.8 2.4T A95B's context window?
Qwen3.8 2.4T A95B supports a context window of 1.0M tokens.
Is Qwen3.8 2.4T A95B worth it in 2026?
Qwen3.8 2.4T A95B is Qwen's frontier tier model. Use the AI Intelligence Calculator to see how its value compares to every other model at your budget.