Clinical Benchmarks

Claude Sonnet 4.5

Released 29 Sep 2025$3 input, $15 output per million tokensproprietaryAlso written as claude-sonnet-4-5-20250929, claude-sonnet-4-5-20250929-thinking, anthropic/claude-sonnet-4-5-20250929-thinking, anthropic/claude-sonnet-4-5-20250929, Claude Sonnet 4.5 (Nonthinking), Claude Sonnet 4.5 (Thinking)Compare with other models

Clinical Benchmarks Index
60.3rank 43 of 148; 73.9 × 0.816 = 60.3, from 2 of 10 boards
Boards
2 of 13
Results
4
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens=30000
    84.52
    Rank 31 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedScribe (Vals AI)
    model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_output_tokens=30000
    84.10
    Rank 36 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  3. MedCode (Vals AI)
    model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_output_tokens=30000
    44.13
    Rank 39 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026
  4. MedCode (Vals AI)
    model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens=30000
    40.57
    Rank 59 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 31MedScribe (Vals AI) model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens… 84.52
    Printed as 84.52%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 31 of 105 (Claude Sonnet 4.5 (Nonthinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["anthropic/claude-sonnet-4-5-20250929"].
    31 | Claude Sonnet 4.5 (Nonthinking) | 84.52%±1.93 | $3/$15 | 44.42s
    Every result from this document
  2. 36MedScribe (Vals AI) model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_outp… 84.10
    Printed as 84.10%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 36 of 105 (Claude Sonnet 4.5 (Thinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["anthropic/claude-sonnet-4-5-20250929-thinking"].
    36 | Claude Sonnet 4.5 (Thinking) | 84.10%±1.87 | $3/$15 | 67.92s
    Every result from this document
  3. 39MedCode (Vals AI) model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_outp… 44.13
    Printed as 44.13%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-sonnet-4-5-20250929-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 39 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["anthropic/claude-sonnet-4-5-20250929-thinking"].
    39 | Claude Sonnet 4.5 (Thinking) | 44.13%±2.00 | $3/$15 | 74.33s
    Every result from this document
  4. 59MedCode (Vals AI) model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens… 40.57
    Printed as 40.57%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-sonnet-4-5-20250929; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 59 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["anthropic/claude-sonnet-4-5-20250929"].
    59 | Claude Sonnet 4.5 (Nonthinking) | 40.57%±2.00 | $3/$15 | 12.01s
    Every result from this document

Other Anthropic models: Claude 3.7 Sonnet, Claude Fable 5, Claude Fable 5.1, Claude Haiku 4.5, Claude Opus 4.1, Claude Opus 4.5, Claude Opus 4.6, Claude Opus 4.7, Claude Opus 4.8, Claude Opus 5, Claude Opus 5.5, Claude Sonnet 4, Claude Sonnet 4.6, Claude Sonnet 5, Claude Sonnet 5.5