Clinical Benchmarks

GLM 5.3

Released 18 Aug 2026$1.4 input, $4.4 output per million tokens1M contextopenAlso written as zai/glm-5.3Compare with other models

Clinical Benchmarks Index
61.1rank 41 of 148; 74.9 × 0.816 = 61.1, from 2 of 10 boards
Boards
2 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID zai/glm-5.3; temperature=1; top_p=0.95; max_output_tokens=30000; reasoning_effort=max
    88.81
    Rank 9 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID zai/glm-5.3; reasoning_effort=max; temperature=1; top_p=0.95; max_output_tokens=30000
    42.86
    Rank 46 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 9MedScribe (Vals AI) model ID zai/glm-5.3; temperature=1; top_p=0.95; max_output_tokens=30000; reaso… 88.81
    Printed as 88.81%Official leaderboard, measured Sep 2026Configuration: model ID zai/glm-5.3; temperature=1; top_p=0.95; max_output_tokens=30000; reasoning_effort=max
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 9 of 105 (GLM 5.3), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["zai/glm-5.3"].
    9 | GLM 5.3 | 88.81%±2.00 | $1.4/$4.4 | 2m02s
    Every result from this document
  2. 46MedCode (Vals AI) model ID zai/glm-5.3; reasoning_effort=max; temperature=1; top_p=0.95; max_outp… 42.86
    Printed as 42.86%Official leaderboard, measured Sep 2026Configuration: model ID zai/glm-5.3; reasoning_effort=max; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 46 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["zai/glm-5.3"].
    46 | GLM 5.3 | 42.86%±2.11 | $1.4/$4.4 | 2m49s
    Every result from this document

Other Zhipu models: GLM 4.7, GLM 5.1, GLM 5.2, GLM 5.3 Flash