Models / DeepSeek
Chat
Reasoning

DeepSeek R1 Distilled Qwen 14B

Qwen 14B distilled with reasoning capabilities from Deepseek R1. Outperforms GPT-4o in math & matches o1-mini on coding.

About model

DeepSeek R1 Distilled Qwen 14B excels at complex problem-solving and reasoning tasks. It leverages large-scale reinforcement learning to develop chain-of-thought capabilities. Suitable for researchers and developers seeking advanced reasoning models.

Performance benchmarks

Model

FrontierMath Tier 4

GPQA Diamond

HLE

SciCode

GDPval-AA

Terminal-Bench 2.1

Agent Arena

FrontierCode

DeepSWE

65.0%

Related open-source models

Competitor closed-source models

Claude Fable 5

87.8%

92.6%

53%

60%

62%

85%

53.5%

70%

Claude Opus 5

73.2%

93.2%

53%

56%

68%

89%

+16.4pp

53.4%

74%

GPT-5.6 Sol

82.9%

94.1%

47%

56%

61%

88%

47.5%

73%

Grok 4.5

24.4%

93.1%

40%

54%

51%

82%

+4.2pp

42.4%

54%

GPT-5.6 Luna

61.0%

91.1%

37%

53%

54%

81%

39.8%

67%

    Related models
    • Model provider
      DeepSeek
    • Type
      Chat
      Reasoning
    • Main use cases
      Chat
      Reasoning
    • Features
      JSON Mode
    • Fine tuning
      Supported
    • Deployment
      Dedicated
    • Parameters
      14.8B
    • Input price

      $0.18 / 1M tokens

    • Output price

      $0.18 / 1M tokens

    • Input modalities
      Text
    • Output modalities
      Text
    • Released
      January 20, 2025
    • Last updated
      September 9, 2025
    • Quantization level
      FP16
    • External link
    • Category
      Chat