Clinical Benchmarks

Claude Haiku 4.5

Released 15 Oct 2025$1 input, $5 output per million tokens200K contextproprietaryAlso written as claude-haiku-4-5, Claude Haiku 4.5 (Thinking), claude-haiku-4-5-20251001-thinking, anthropic/claude-haiku-4-5-20251001-thinkingCompare with other models

Clinical Benchmarks Index
50.0rank 66 of 148; 61.3 × 0.816 = 50.0, from 2 of 10 boards
Boards
3 of 13
Results
3
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_output_tokens=30000
    85.23
    Rank 28 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_output_tokens=30000
    32.68
    Rank 84 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

EHR and workflow agents

  1. CHI-Bench
    Listed as claude-code + claude-haiku-4-5
    All Domains pass@1; claude-code harness
    6.2
    Rank 36 of 43 here, 45 models on the boardLeader erius + claude-opus-5 54.7
    Official leaderboard
    Measured Apr 2026

Sources

Open a line for the quote and page.

  1. 28MedScribe (Vals AI) model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_outpu… 85.23
    Printed as 85.23%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 29 of 105 (Claude Haiku 4.5 (Thinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["anthropic/claude-haiku-4-5-20251001-thinking"].
    29 | Claude Haiku 4.5 (Thinking) | 85.23%±1.90 | $1/$5 | 66.20s
    Every result from this document
  2. 84MedCode (Vals AI) model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_outpu… 32.68
    Printed as 32.68%Official leaderboard, measured Sep 2026Configuration: model ID anthropic/claude-haiku-4-5-20251001-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 84 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["anthropic/claude-haiku-4-5-20251001-thinking"].
    84 | Claude Haiku 4.5 (Thinking) | 32.68%±2.00 | $1/$5 | 34.29s
    Every result from this document
  3. 36CHI-Bench All Domains pass@1; claude-code harness 6.2
    Printed as 6.2%Official leaderboard, measured Apr 2026Configuration: All Domains pass@1; claude-code harness
    CHI-Bench leaderboard (actAVA) official leaderboard, actAVA, 12 Aug 2026. CHI-Bench v1.0.0, All Domains, rank 37, Agent/Model and Accuracy columns; PA/UM/CM follow; board last updated 2026-08-12; board Date column 2026-05-01; accessed 2026-09-30. Linked submission provenance: run 2026-04-29T19:36:29Z to 2026-04-30T19:34:41Z; submitted_at 2026-04-30T19:34:41Z.
    37 | claude-code | claude-haiku-4-5 | Proprietary | 6.2% | 0.0% | 14.7% | 4.0% | 2026-05-01
    Every result from this document

Other Anthropic models: Claude 3.7 Sonnet, Claude Fable 5, Claude Fable 5.1, Claude Opus 4.1, Claude Opus 4.5, Claude Opus 4.6, Claude Opus 4.7, Claude Opus 4.8, Claude Opus 5, Claude Opus 5.5, Claude Sonnet 4, Claude Sonnet 4.5, Claude Sonnet 4.6, Claude Sonnet 5, Claude Sonnet 5.5