Qwen3 VL 235B pricing

Alibaba (Qwen) · dashscope-qwen3-vl-235b-a22b-instruct · 131K context · 18th cheapest of 64 current models

This row is flagged for review: OpenRouter disagrees (48%): $0.21/$1.9 vs $0.4/$1.6. Check the vendor’s own page before relying on it.

What Qwen3 VL 235B costs

USD, standard tier, global endpoint
RatePer 1M tokensPer 1K tokensWhat it covers
Input $0.4 $0.0004 Every token you send: prompt, history, documents.
Output $1.6 $0.0016 Every token the model generates, including hidden reasoning tokens.

What that means in practice

The same three workloads are costed on every model page, so these numbers are directly comparable across the catalog. No prompt caching and no batch discount is assumed — the figures are what you pay before you optimise anything.

Support chatbot

2,500 six-turn conversations a day

$1,449 / month

$0.019 per conversation · $17,630 a year

An 800-token system prompt, 400-token questions and 900-token answers, over six turns — each turn re-sending everything before it.

Document summariser

1,000 single-shot summaries a day

$185 / month

$0.0062 per conversation · $2,248 a year

One 12,000-token document in, an 800-token summary out, no conversation history. The case where input dominates the bill.

Coding agent

200 twelve-turn sessions a day

$1,051 / month

$0.175 per conversation · $12,790 a year

A 3,000-token system prompt, 1,500-token instructions and 2,500-token responses over twelve turns. Long sessions are where the quadratic cost of re-sent history stops being theoretical.

Change the assumptions in the calculator

Cheaper models in the catalog

Nothing cheaper is scored as highly as Qwen3 VL 235B, so these are simply the cheaper options, best-scored first. Trading down here is a real trade.

ModelInputOutputBlended saving
DeepSeek V4 Pro $0.435 $0.87 22%
DeepSeek V3.2 $0.28 $0.4 56%
GPT-5.6 Luna $0.2 $1.2 36%
MiniMax M3 $0.3 $1.2 25%
Qwen3 Next 80B $0.15 $1.2 41%

Specification

Context window131K tokens
Maximum output33K tokens
ProviderAlibaba (Qwen)
StatusCurrent
Reasoning modelYes
VisionNo
Token countingEstimated at ~3.5 characters per token (Estimate — Qwen tokenizer, calibrated for mixed English/Chinese)

Where these numbers come from

Last checked against its source on 2026-08-02, and the published rates last actually moved on 2026-08-02. Source of record: LiteLLM’s community-maintained price map.

Alibaba (Qwen)’s own pricing page

Regional and data-residency premiums, priority tiers, server-side tool fees and negotiated discounts are not included. Prices change without notice; this page is rebuilt every time the catalog does.