Clinical Benchmarks

Artificial Analysis Healthcare & Medical Index: current results

Artificial Analysis · 27 of 159 models scored; composite of 4 underlying benchmarks · index updated August 16, 2026

Claude Opus 5 (Adaptive Reasoning, Max Effort) holds the top current result on Artificial Analysis Healthcare & Medical Index, 51 as of 2026-08, per Artificial Analysis healthcare capability page. Weighted composite for healthcare and medical work: Medical & Health Knowledge 35%, Agentic Knowledge Work 25%, Non-Hallucination 15%, Reasoning 15%, Agentic Customer Interaction 10%, drawn from AA-Omniscience, GDPval-AA v2, Humanity's Last Exam, and tau3-Banking.

Current results

Result detail

#modelscoreas of
1Anthropic logoClaude Opus 5 (Adaptive Reasoning, Max Effort) Anthropic
all underlying benchmarks run independently by Artificial Analysis
512026-08
2Anthropic logoClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) Anthropic512026-08
3Anthropic logoClaude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback) Anthropic512026-08

Scores appear exactly as Artificial Analysis healthcare capability page publishes them (independently run). A weighted composite of general-purpose benchmarks tilted toward health-relevant slices rather than purpose-built clinical tasks, which is why this index files it as a capability index. The page differentiates only its top scores numerically.

About the benchmark

publisherArtificial Analysis
categorycomposite indices
released2026-08
size27 of 159 models scored; composite of 4 underlying benchmarks
scaleindex score, higher better
result basisindependently run
sourceArtificial Analysis healthcare capability page
last frontier result2026-08

What is Artificial Analysis Healthcare & Medical Index?

Artificial Analysis Healthcare & Medical Index is a composite benchmark from Artificial Analysis, released 2026-08: 27 of 159 models scored; composite of 4 underlying benchmarks, scored on a index score scale. Weighted composite for healthcare and medical work: Medical & Health Knowledge 35%, Agentic Knowledge Work 25%, Non-Hallucination 15%, Reasoning 15%, Agentic Customer Interaction 10%, drawn from AA-Omniscience, GDPval-AA v2, Humanity's Last Exam, and tau3-Banking.

Which model leads Artificial Analysis Healthcare & Medical Index?

Claude Opus 5 (Adaptive Reasoning, Max Effort) (Anthropic) holds the top current result on Artificial Analysis Healthcare & Medical Index at 51, per Artificial Analysis healthcare capability page, as of 2026-08.

Where do the Artificial Analysis Healthcare & Medical Index numbers come from?

From Artificial Analysis healthcare capability page (independently run). A weighted composite of general-purpose benchmarks tilted toward health-relevant slices rather than purpose-built clinical tasks, which is why this index files it as a capability index. The page differentiates only its top scores numerically.

The rest of the field is on the index, and how sources qualify is on the methodology page. Model names in the table link to cross-benchmark pages.