DeepSeek
Available in Shortly AI
DeepSeek V3 0324
Released December 26, 2024
Tool Use
DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.
- Context window
- 163.8K tokens
- Maximum output
- 163.8K tokens
- Input pricing
- $0.27/M
- Output pricing
- $1.12/M
Independent benchmark
LiveBench historical snapshot
62.8overall
Historical benchmark data from LiveBench, release 2025-04-25. Scores are pinned and are not fetched on page load.
44.3
68.9
71.4
64.0
46.8
81.5
| Category | Score |
|---|---|
| Reasoning | 44.278 |
| Coding | 68.907 |
| Mathematics | 71.437 |
| Data Analysis | 64.019 |
| Language | 46.823 |
| Instruction Following | 81.471 |
Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.
FAQ
Similar AI models
Qwen 3.7 FlashAlibabaThe Qwen3.7 native vision-language Flash model series delivers a comprehensive upgrade over 3.6-Flash in multimodal understanding and agent execution. This model particularly excels in enhanced multimodal foundations with stronger universal object recognition, further improved real-world perception and spatial intelligence, significantly upgraded multimodal agent capabilities for Search Agent and CI Agent scenarios with more stable end-to-end task execution, as well as optimized multimodal coding for a smoother vibe coding experience.GPT 5.6 LunaOpenAIGPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.Hy3TencentTencent Hy3 is an open-source Mixture-of-Experts (MoE) large language model developed by Tencent's Hunyuan team. It has 295B total parameters, 21B active parameters, a 3.8B-parameter MTP layer, and a 256K context window. Hy3 is designed for agentic workflows, coding, reasoning, long-context understanding, tool-use scenarios, document analysis, and structured generation.
The model improves on Hy3 Preview with stronger tool-call reliability, better output-format stability, improved long-context retention, and lower hallucination rates in Tencent's internal evaluations.
Available with one subscription
Start with DeepSeek V3 0324 in Shortly AI
Keep your conversations, files and model choices together in one workspace.