Models / ZAI
Reasoning
Chat

GLM-5.1

Refined post-training for coding and agentic engineering workflows

This model is not available on Together’s Serverless API.

Performance benchmarks

Model

FrontierMath Tier 4

GPQA Diamond

HLE

SciCode

GDPval-AA

Terminal-Bench 2.1

Agent Arena

FrontierCode

DeepSWE

86.8%

28%

44%

38%

62%

+1.2pp

Related open-source models

Competitor closed-source models

Claude Fable 5

87.8%

92.6%

53%

60%

62%

85%

53.5%

70%

Claude Opus 5

73.2%

93.2%

53%

56%

68%

89%

+16.4pp

53.4%

74%

GPT-5.6 Sol

82.9%

94.1%

47%

56%

61%

88%

47.5%

73%

Grok 4.5

24.4%

93.1%

40%

54%

51%

82%

+4.2pp

42.4%

54%

GPT-5.6 Luna

61.0%

91.1%

37%

53%

54%

81%

39.8%

67%

Related models
  • Model provider
    ZAI
  • Type
    Reasoning
    Chat
  • Main use cases
    Reasoning
  • Features
    Function Calling
    JSON Mode
  • Parameters
    754B
  • Activated parameters
    40B
  • Context length
    200K
  • Input price

    $1.40 / 1M tokens

    $0.26 (cached)/1M

  • Output price

    $4.40 / 1M tokens

  • Input modalities
    Text
  • Output modalities
    Text
  • Released
    March 26, 2026
  • Quantization level
    FP4
  • Category
    Chat