Moonshot AI
Permanent model reference
Kimi K2 Thinking
Released November 6, 2025
Kimi K2 Thinking is an advanced open-source thinking model by Moonshot AI. It can execute up to 200 – 300 sequential tool calls without human interference, reasoning coherently across hundreds of steps to solve complex problems. Built as a thinking agent, it reasons step by step while using tools, achieving state-of-the-art performance on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks, with major gains in reasoning, agentic search, coding, writing, and general capabilities.
- Context window
- 216.1K tokens
- Maximum output
- 216.1K tokens
- Input pricing
- $0.47/M
- Output pricing
- $2.00/M
Independent benchmark
LiveBench historical snapshot
Historical benchmark data from LiveBench, release 2025-12-23. Scores are pinned and are not fetched on page load.
| Category | Score |
|---|---|
| Reasoning | 63.486 |
| Coding | 67.437 |
| Agentic Coding | 38.333 |
| Mathematics | 90.794 |
| Data Analysis | 70.568 |
| Language | 66.453 |
| Instruction Following | 62.034 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
One subscription, 100+ models
Find your next model in Shortly AI
Keep your conversations, files and model choices together in one workspace.