Alibaba
Available in Shortly AI
Qwen 3 32B
Released April 28, 2025
ReasoningTool Use
Qwen3-32B is a world-class model with comparable quality to DeepSeek R1 while outperforming GPT-4.1 and Claude Sonnet 3.7. It excels in code-gen, tool-calling, and advanced reasoning, making it an exceptional model for a wide range of production use cases.
- Context window
- 128K tokens
- Maximum output
- 8.2K tokens
- Input pricing
- $0.16/M
- Output pricing
- $0.64/M
Independent benchmark
LiveBench historical snapshot
43.6overall
Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.
48.3
66.0
3.3
67.4
46.5
55.5
17.8
| Category | Score |
|---|---|
| Reasoning | 48.255 |
| Coding | 66.029 |
| Agentic Coding | 3.333 |
| Mathematics | 67.440 |
| Data Analysis | 46.540 |
| Language | 55.542 |
| Instruction Following | 17.771 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
GLM 5.3 FlashZ.aiGLM-5.3-Flash is Z.ai’s native multimodal coding model, featuring 320B total parameters, 18B activated parameters, and a 1M-token context window. Its efficient hybrid attention architecture supports visual coding, tool use, and end-to-end professional workflows across code, browsers, documents, and graphical interfaces.DeepSeek V4 Flash Vision ExpDeepSeekDeepSeek-V4-Flash-Vision-Exp is the first experimental multimodal model in the DeepSeek-V4 family. It builds on DeepSeek-V4-Flash by adding visual modules and continued training for visual understanding, with substantially improved multimodal agent capabilities while remaining comparable on text-only agent tasks.Nemotron 3.5 Lightning 30BNVIDIANVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 is a large language model (LLM) trained by NVIDIA.
The model employs a hybrid Mixture-of-Experts architecture, utilizing interleaved Mamba-2 and MoE layers, along with select Attention layers. The Lightning 3.5 model is released alongside a number of speculative decoding methods for faster text generation. The model has 3B active parameters and 30B parameters in total.
Available with one subscription
Start with Qwen 3 32B in Shortly AI
Keep your conversations, files and model choices together in one workspace.