DeepSeek
Permanent model reference
DeepSeek V3 0324
Released December 26, 2024
Tool Use
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.
- Context window
- 163.8K tokens
- Maximum output
- 163.8K tokens
- Input pricing
- $0.27/M
- Output pricing
- $1.12/M
Independent benchmark
LiveBench historical snapshot
62.8overall
Historical benchmark data from LiveBench, release 2025-04-25. Scores are pinned and are not fetched on page load.
44.3
68.9
71.4
64.0
46.8
81.5
| Category | Score |
|---|---|
| Reasoning | 44.278 |
| Coding | 68.907 |
| Mathematics | 71.437 |
| Data Analysis | 64.019 |
| Language | 46.823 |
| Instruction Following | 81.471 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.DeepSeek V4 FlashDeepSeekDeepseek V4 Flash is an AI model available through Shortly AI.
One subscription, 100+ models
Find your next model in Shortly AI
Keep your conversations, files and model choices together in one workspace.