Models / DeepSeek
Chat
Reasoning

DeepSeek R1 Distilled Llama 70B

Llama 70B distilled with reasoning capabilities from Deepseek R1. Surpasses GPT-4o with 94.5% on MATH-500 & matches o1-mini on coding.

About model

DeepSeek R1 Distilled Llama 70B performs complex reasoning tasks, excelling in math, code, and reasoning benchmarks. It is designed for researchers and developers seeking advanced language models.

Performance benchmarks

Model

FrontierMath Tier 4

GPQA Diamond

HLE

SciCode

GDPval-AA

Terminal-Bench 2.1

Agent Arena

FrontierCode

DeepSWE

65.0%

Related open-source models

Competitor closed-source models

Claude Fable 5

87.8%

92.6%

53%

60%

62%

85%

53.5%

70%

Claude Opus 5

73.2%

93.2%

53%

56%

68%

89%

+16.4pp

53.4%

74%

GPT-5.6 Sol

82.9%

94.1%

47%

56%

61%

88%

47.5%

73%

Grok 4.5

24.4%

93.1%

40%

54%

51%

82%

+4.2pp

42.4%

54%

GPT-5.6 Luna

61.0%

91.1%

37%

53%

54%

81%

39.8%

67%

    Related models
    • Model provider
      DeepSeek
    • Type
      Chat
      Reasoning
    • Main use cases
      Chat
      Reasoning
    • Features
      JSON Mode
    • Fine tuning
      Supported
    • Deployment
      Dedicated
    • Parameters
      70B
    • Context length
      128K
    • Input price

      $2.00 / 1M tokens

    • Output price

      $2.00 / 1M tokens

    • Input modalities
      Text
    • Output modalities
      Text
    • Released
      January 20, 2025
    • Last updated
      December 22, 2025
    • Quantization level
      FP16
    • External link
    • Category
      Chat