Alibaba
Available in Shortly AI
Qwen3-30B-A3B
Released April 28, 2025
ReasoningTool Use
Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support
- Context window
- 41K tokens
- Maximum output
- 16.4K tokens
- Input pricing
- $0.12/M
- Output pricing
- $0.50/M
Independent benchmark
LiveBench historical snapshot
39.0overall
Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.
36.7
48.9
1.7
65.3
44.9
54.5
21.1
| Category | Score |
|---|---|
| Reasoning | 36.678 |
| Coding | 48.883 |
| Agentic Coding | 1.667 |
| Mathematics | 65.347 |
| Data Analysis | 44.923 |
| Language | 54.465 |
| Instruction Following | 21.108 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
GLM 5.3 FlashZ.aiGLM-5.3-Flash is Z.ai’s native multimodal coding model, featuring 320B total parameters, 18B activated parameters, and a 1M-token context window. Its efficient hybrid attention architecture supports visual coding, tool use, and end-to-end professional workflows across code, browsers, documents, and graphical interfaces.DeepSeek V4 Flash Vision ExpDeepSeekDeepSeek-V4-Flash-Vision-Exp is the first experimental multimodal model in the DeepSeek-V4 family. It builds on DeepSeek-V4-Flash by adding visual modules and continued training for visual understanding, with substantially improved multimodal agent capabilities while remaining comparable on text-only agent tasks.Nemotron 3.5 Lightning 30BNVIDIANVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4 is a large language model (LLM) trained by NVIDIA.
The model employs a hybrid Mixture-of-Experts architecture, utilizing interleaved Mamba-2 and MoE layers, along with select Attention layers. The Lightning 3.5 model is released alongside a number of speculative decoding methods for faster text generation. The model has 3B active parameters and 30B parameters in total.
Available with one subscription
Start with Qwen3-30B-A3B in Shortly AI
Keep your conversations, files and model choices together in one workspace.