Alibaba
Available in Shortly AI
Qwen3 Next 80B A3B Instruct
Released September 11, 2025
Tool Use
A new generation of open-source, non-thinking mode model powered by Qwen3. This version demonstrates superior Chinese text understanding, augmented logical reasoning, and enhanced capabilities in text generation tasks over the previous iteration (Qwen3-235B-A22B-Instruct-2507).
- Context window
- 262.1K tokens
- Maximum output
- 262.1K tokens
- Input pricing
- $0.15/M
- Output pricing
- $1.20/M
Independent benchmark
LiveBench historical snapshot
53.1overall
Historical benchmark data from LiveBench, release 2025-12-23. Scores are pinned and are not fetched on page load.
54.7
68.2
10.0
84.9
68.6
66.3
19.2
| Category | Score |
|---|---|
| Reasoning | 54.745 |
| Coding | 68.203 |
| Agentic Coding | 10.000 |
| Mathematics | 84.904 |
| Data Analysis | 68.627 |
| Language | 66.338 |
| Instruction Following | 19.188 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
GLM 5.3 FlashZ.aiGLM-5.3-Flash is Z.ai’s native multimodal coding model, featuring 320B total parameters, 18B activated parameters, and a 1M-token context window. Its efficient hybrid attention architecture supports visual coding, tool use, and end-to-end professional workflows across code, browsers, documents, and graphical interfaces.DeepSeek V4 Flash Vision ExpDeepSeekDeepSeek-V4-Flash-Vision-Exp is the first experimental multimodal model in the DeepSeek-V4 family. It builds on DeepSeek-V4-Flash by adding visual modules and continued training for visual understanding, with substantially improved multimodal agent capabilities while remaining comparable on text-only agent tasks.Nemotron 3.5 Lightning 30BNVIDIANVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 is a large language model (LLM) trained by NVIDIA.
The model employs a hybrid Mixture-of-Experts architecture, utilizing interleaved Mamba-2 and MoE layers, along with select Attention layers. The Lightning 3.5 model is released alongside a number of speculative decoding methods for faster text generation. The model has 3B active parameters and 30B parameters in total.
Available with one subscription
Start with Qwen3 Next 80B A3B Instruct in Shortly AI
Keep your conversations, files and model choices together in one workspace.