QwenUpdated 18h ago

Qwen: Qwen3 Next 80B A3B Thinking API Pricing

Live token cost for Qwen: Qwen3 Next 80B A3B Thinking from Qwen. Use the figures below for budgeting, then tune your exact request mix in the interactive calculator. Prices refresh every 24 hours from OpenRouter.

Input

$0.150

/ 1M tokens

Output

$1.20

/ 1M tokens

Cached input

—

Not supported

Capabilities

262K context33K max outputReasoning tokens

Qwen: Qwen3 Next 80B A3B Thinking cost at scale

Estimated monthly cost across common production volumes. Assumes 30-day months and the request shapes shown.

Tier	Requests / day	In / out tokens	$ / month
Hobby	1,000	500 / 200	$9.45
Startup	10,000	1,500 / 500	$247.50
Growth	100,000	3,000 / 800	$4,230.00
Enterprise	1,000,000	8,000 / 2,000	$108,000.00

Open Qwen: Qwen3 Next 80B A3B Thinking in interactive calculator →

Compare Qwen: Qwen3 Next 80B A3B Thinking vs.

OpenAI

OpenAI: GPT-4o

$2.50 in · $10.00 out

Compare side-by-side →

Anthropic

Anthropic: Claude Haiku 4.5

$1.00 in · $5.00 out

Compare side-by-side →

Google

Google: Gemini 2.5 Pro

$1.25 in · $10.00 out

Compare side-by-side →

Frequently asked questions

How much does Qwen: Qwen3 Next 80B A3B Thinking cost?

Qwen: Qwen3 Next 80B A3B Thinking costs $0.15 per 1M input tokens and $1.20 per 1M output tokens. A typical 1,500-token in / 500-token out request costs $0.00082.

Does Qwen: Qwen3 Next 80B A3B Thinking support cached input?

No. Qwen: Qwen3 Next 80B A3B Thinking does not currently expose cached-input pricing through Qwen. Every input token is billed at the full rate.

What is the Qwen: Qwen3 Next 80B A3B Thinking context window?

Qwen: Qwen3 Next 80B A3B Thinking supports a context window of 262,144 tokens (262K). Max output per response is 32,768 tokens.

What is Qwen: Qwen3 Next 80B A3B Thinking good for?

Qwen: Qwen3 Next 80B A3B Thinking is a good fit for general-purpose LLM tasks via the Qwen API — chat, code, writing, summarization. For other use cases, run your specific input/output mix through the interactive calculator to compare against alternative models.