Toggle navigation

Alibaba

Available in Shortly AI

Qwen3 Next 80B A3B Thinking

Released September 11, 2025

ReasoningTool Use

Qwen3-Next uses a highly sparse MoE design: 80B total parameters, but only ~3B activated per inference step. Experiments show that, with global load balancing, increasing total expert parameters while keeping activated experts fixed steadily reduces training loss.Compared to Qwen3’s MoE (128 total experts, 8 routed), Qwen3-Next expands to 512 total experts, combining 10 routed experts + 1 shared expert — maximizing resource usage without hurting performance. The Qwen3-Next-80B-A3B-Thinking excels at complex reasoning tasks — outperforming higher-cost models like Qwen3-30B-A3B-Thinking-2507 and Qwen3-32B-Thinking, outpeforming the closed-source Gemini-2.5-Flash-Thinking on multiple benchmarks, and approaching the performance of our top-tier model Qwen3-235B-A22B-Thinking-2507.

Context window
262.1K tokens
Maximum output
262.1K tokens
Input pricing
$0.15/M
Output pricing
$1.20/M

Independent benchmark

LiveBench historical snapshot

54.9overall

Historical benchmark data from LiveBench, release 2025-12-23. Scores are pinned and are not fetched on page load.

Qwen3 Next 80B A3B Thinking LiveBench category scores
CategoryScore
Reasoning58.164
Coding60.656
Agentic Coding8.333
Mathematics86.347
Data Analysis73.164
Language56.312
Instruction Following41.542

Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.

FAQ

Similar AI models

Available with one subscription

Start with Qwen3 Next 80B A3B Thinking in Shortly AI

Keep your conversations, files and model choices together in one workspace.