Meta
Available in Shortly AI
Llama 4 Maverick 17B Instruct
Released April 5, 2025
Vision (Image)Tool Use
As a general purpose LLM, Llama 4 Maverick contains 17 billion active parameters, 128 experts, and 400 billion total parameters, offering high quality at a lower price compared to Llama 3.3 70B.
- Context window
- 128K tokens
- Maximum output
- 8.2K tokens
- Input pricing
- $0.24/M
- Output pricing
- $0.97/M
Independent benchmark
LiveBench historical snapshot
54.4overall
Historical benchmark data from LiveBench, release 2025-04-02. Scores are pinned and are not fetched on page load.
43.8
37.4
60.6
59.0
49.6
75.7
| Category | Score |
|---|---|
| Reasoning | 43.833 |
| Coding | 37.429 |
| Mathematics | 60.579 |
| Data Analysis | 59.033 |
| Language | 49.648 |
| Instruction Following | 75.746 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Qwen 3.7 FlashAlibabaThe Qwen3.7 native vision-language Flash model series delivers a comprehensive upgrade over 3.6-Flash in multimodal understanding and agent execution. This model particularly excels in enhanced multimodal foundations with stronger universal object recognition, further improved real-world perception and spatial intelligence, significantly upgraded multimodal agent capabilities for Search Agent and CI Agent scenarios with more stable end-to-end task execution, as well as optimized multimodal coding for a smoother vibe coding experience.GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.
Available with one subscription
Start with Llama 4 Maverick 17B Instruct in Shortly AI
Keep your conversations, files and model choices together in one workspace.