Qwen3 Next 80B pricing
Alibaba (Qwen) ·
dashscope-qwen3-next-80b-a3b-instruct ·
262K context
· 5th cheapest of 64 current models
This row is flagged for review: OpenRouter disagrees (33%): $0.1/$1.1 vs $0.15/$1.2. Check the vendor’s own page before relying on it.
What Qwen3 Next 80B costs
| Rate | Per 1M tokens | Per 1K tokens | What it covers |
|---|---|---|---|
| Input | $0.15 | $0.00015 | Every token you send: prompt, history, documents. |
| Output | $1.2 | $0.0012 | Every token the model generates, including hidden reasoning tokens. |
What that means in practice
The same three workloads are costed on every model page, so these numbers are directly comparable across the catalog. No prompt caching and no batch discount is assumed — the figures are what you pay before you optimise anything.
Support chatbot
2,500 six-turn conversations a day
$786 / month
$0.010 per conversation · $9,568 a year
An 800-token system prompt, 400-token questions and 900-token answers, over six turns — each turn re-sending everything before it.
Document summariser
1,000 single-shot summaries a day
$83.70 / month
$0.0028 per conversation · $1,018 a year
One 12,000-token document in, an 800-token summary out, no conversation history. The case where input dominates the bill.
Coding agent
200 twelve-turn sessions a day
$502 / month
$0.084 per conversation · $6,110 a year
A 3,000-token system prompt, 1,500-token instructions and 2,500-token responses over twelve turns. Long sessions are where the quadratic cost of re-sent history stops being theoretical.
Change the assumptions in the calculator
Cheaper, and scored at least as capable
Every model below costs less on a blended rate and carries a capability score at least as high as Qwen3 Next 80B’s. The score is a rough estimate, not a benchmark — treat it as a shortlist, not a verdict.
| Model | Input | Output | Blended saving | |
|---|---|---|---|---|
| DeepSeek V3.2 | $0.28 | $0.4 | 25% |
Head to head
- Qwen3 Next 80B vs Nova 2 Lite
- Qwen3 Next 80B vs DeepSeek V4 Flash
- Qwen3 Next 80B vs Gemini 3.5 Flash-Lite
- Qwen3 Next 80B vs GPT-5.6 Luna
- Qwen3 Next 80B vs Mistral Large 3
- Qwen3 Next 80B vs Grok 4.1 Fast
Specification
| Context window | 262K tokens |
|---|---|
| Maximum output | 66K tokens |
| Provider | Alibaba (Qwen) |
| Status | Current |
| Reasoning model | Yes |
| Vision | No |
| Token counting | Estimated at ~3.5 characters per token (Estimate — Qwen tokenizer, calibrated for mixed English/Chinese) |
Where these numbers come from
Last checked against its source on 2026-08-02, and the published rates last actually moved on 2026-08-02. Source of record: LiteLLM’s community-maintained price map.
Alibaba (Qwen)’s own pricing page
Regional and data-residency premiums, priority tiers, server-side tool fees and negotiated discounts are not included. Prices change without notice; this page is rebuilt every time the catalog does.