Clinical Benchmarks

Claude 3.7 Sonnet

Released 24 Feb 2025proprietaryCompare with other models

Clinical Benchmarks Index
20.1rank 126 of 148; 34.8 × 0.577 = 20.1, from 1 of 10 boards
Boards
1 of 13
Results
1
Latest measurement
May 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Clinical reasoning and knowledge

  1. 0.450
    Rank 9 of 10 here, 11 models on the boardLeader Gemini 3.1 Pro (Preview) 0.652
    Official leaderboard
    Measured May 2026

Sources

Open a line for the quote and page.

  1. 9MedHELM 0.450
    Printed as 0.45Official leaderboard, measured May 2026
    MedHELM leaderboard (medhelm.org), v5.0.0 official leaderboard, Stanford CRFM (MedHELM), 14 May 2026. medhelm.org home, 'Current leaders Mean win rate v5.0.0' table ('10 of 11 models · Updated 14 May 2026')
    9 Claude 3.7 Sonnet (20250219) Anthropic 0.45
    Every result from this document

Other Anthropic models: Claude Fable 5, Claude Fable 5.1, Claude Haiku 4.5, Claude Opus 4.1, Claude Opus 4.5, Claude Opus 4.6, Claude Opus 4.7, Claude Opus 4.8, Claude Opus 5, Claude Opus 5.5, Claude Sonnet 4, Claude Sonnet 4.5, Claude Sonnet 4.6, Claude Sonnet 5, Claude Sonnet 5.5