DeepSeek-V3-0324
Mixture-of-Experts model challenging top AI models at much lower cost. Updated on March 24th, 2025.
About model
DeepSeek-V3-0324 is a strong Mixture-of-Experts language model with 671B parameters, 37B activated per token, designed for efficient inference and cost-effective training. It excels in performance, outpacing other open-source models and rivaling leading closed-source models. Suitable for applications requiring high-quality language understanding and generation.
This endpoint was updated on March 24th, 2025 to use the weights of the improved DeepSeek-V3-0324 model.
Model | FrontierMath Tier 4 | GPQA Diamond | HLE | SciCode | GDPval-AA | Terminal-Bench 2.1 | Agent Arena | FrontierCode | DeepSWE |
|---|---|---|---|---|---|---|---|---|---|
DeepSeek-V3-0324 | 35.3% | Related open-source models | Competitor closed-source models | ||||||
87.8% | 92.6% | 53% | 60% | 62% | 85% | 53.5% | 70% | ||
73.2% | 93.2% | 53% | 56% | 68% | 89% | +16.4pp | 53.4% | 74% | |
82.9% | 94.1% | 47% | 56% | 61% | 88% | 47.5% | 73% | ||
24.4% | 93.1% | 40% | 54% | 51% | 82% | +4.2pp | 42.4% | 54% | |
61.0% | 91.1% | 37% | 53% | 54% | 81% | 39.8% | 67% |
- TypeLLMChat
- Main use casesChatFunction Calling
- FeaturesFunction CallingJSON Mode
- Fine tuningSupported
- DeploymentDedicated
- Parameters671B
- Context length131K
- Input price
$1.25 / 1M tokens
- Output price
$1.25 / 1M tokens
- Input modalitiesText
- Output modalitiesText
- ReleasedDecember 25, 2024
- Last updatedJanuary 22, 2026
- Quantization levelFP8
- External link
- CategoryChat
