Anthropic
Available in Shortly AI
Claude Sonnet 4
Released May 22, 2025
ReasoningVision (Image)Tool Use
Claude Sonnet 4 balances impressive performance for coding with the right speed and cost for high-volume use cases: Coding: Handle everyday development tasks with enhanced performance-power code reviews, bug fixes, API integrations, and feature development with immediate feedback loops.
- Context window
- 1M tokens
- Maximum output
- 64K tokens
- Input pricing
- $3.00/M
- Output pricing
- $15.00/M
Independent benchmark
LiveBench historical snapshot
51.0overall
Historical benchmark data from LiveBench, release 2026-01-08. Scores are pinned and are not fetched on page load.
39.7
80.7
38.3
60.4
44.1
71.0
22.7
| Category | Score |
|---|---|
| Reasoning | 39.673 |
| Coding | 80.741 |
| Agentic Coding | 38.333 |
| Mathematics | 60.357 |
| Data Analysis | 44.067 |
| Language | 71.014 |
| Instruction Following | 22.679 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Muse Spark 1.3MetaMuse Spark 1.3 is Metaβs multimodal reasoning model for long-horizon agentic and coding workflows. With a 1M-token context window, reliable tool calling, higher first-attempt accuracy, and native understanding of video, images, and documents, it helps developers build capable coding agents and AI development workflows with fewer unnecessary turns and cleaner output.GLM 5.3Z.aiGLM 5.3 delivers comprehensive advancements in complex software engineering and agent capabilities. It uses the same base model as GLM-5.2, with all improvements driven by post-training.Muse Spark 1.2MetaA coding-optimized model purpose-built for agentic workflows. Improvements to code generation, debugging, and codebase understanding β with a 1M context window that handles your entire project in one session.
Available with one subscription
Start with Claude Sonnet 4 in Shortly AI
Keep your conversations, files and model choices together in one workspace.