Alibaba
Available in Shortly AI
Qwen3 Next 80B A3B Thinking
Released September 11, 2025
Qwen3-Next uses a highly sparse MoE design: 80B total parameters, but only ~3B activated per inference step. Experiments show that, with global load balancing, increasing total expert parameters while keeping activated experts fixed steadily reduces training loss.Compared to Qwen3’s MoE (128 total experts, 8 routed), Qwen3-Next expands to 512 total experts, combining 10 routed experts + 1 shared expert — maximizing resource usage without hurting performance. The Qwen3-Next-80B-A3B-Thinking excels at complex reasoning tasks — outperforming higher-cost models like Qwen3-30B-A3B-Thinking-2507 and Qwen3-32B-Thinking, outpeforming the closed-source Gemini-2.5-Flash-Thinking on multiple benchmarks, and approaching the performance of our top-tier model Qwen3-235B-A22B-Thinking-2507.
- Context window
- 262.1K tokens
- Maximum output
- 262.1K tokens
- Input pricing
- $0.15/M
- Output pricing
- $1.20/M
Independent benchmark
LiveBench historical snapshot
Historical benchmark data from LiveBench, release 2025-12-23. Scores are pinned and are not fetched on page load.
| Category | Score |
|---|---|
| Reasoning | 58.164 |
| Coding | 60.656 |
| Agentic Coding | 8.333 |
| Mathematics | 86.347 |
| Data Analysis | 73.164 |
| Language | 56.312 |
| Instruction Following | 41.542 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Available with one subscription
Start with Qwen3 Next 80B A3B Thinking in Shortly AI
Keep your conversations, files and model choices together in one workspace.