Health Optimization Bench: current results
Health Optimization Bench · 89 released tasks (v1 evidence suite) · index updated August 16, 2026
Claude Fable 5 holds the top current result on Health Optimization Bench, 83.8 as of 2026-08, per healthoptimizationbench.com. Frontier models on hard, freshness-dependent questions in preventive and optimization medicine. Every task is written against a primary source, audited by model families that did not author it, and scored blind by a panel of independent families. The v1 release set covers incretin therapeutics evidence.
This benchmark has a dedicated full leaderboard, with methodology and per-model pages, at healthoptimizationbench.com. The top of its table is mirrored below.
Current top results
- 1
Claude Fable 583.8
- 2
Grok 4.681.3
- 3
Claude Opus 578.3
Result detail
| # | model | score | as of | |
|---|---|---|---|---|
| 1 | Claude Fable 5 Anthropic | 83.8 | 2026-08 | |
| 2 | Grok 4.6 xAI | 81.3 | 2026-08 | |
| 3 | Claude Opus 5 Anthropic | 78.3 | 2026-08 | |
Scores appear exactly as healthoptimizationbench.com publishes them (independently run). Cross-family authoring with blind three-family panel grading; the authoring family never grades its own task, and 95 percent bootstrap confidence intervals accompany every score on the site.
About the benchmark
| publisher | Health Optimization Bench |
|---|---|
| category | rubric-graded benchmarks |
| released | 2026-08 |
| size | 89 released tasks (v1 evidence suite) |
| scale | 0-100 rubric credit, higher better |
| result basis | independently run |
| source | healthoptimizationbench.com |
| last frontier result | 2026-08 |
What is Health Optimization Bench?
Health Optimization Bench is a rubric-graded benchmark from healthoptimizationbench.com, released 2026-08: 89 released tasks (v1 evidence suite), scored on a 0-100 rubric credit scale. Frontier models on hard, freshness-dependent questions in preventive and optimization medicine. Every task is written against a primary source, audited by model families that did not author it, and scored blind by a panel of independent families. The v1 release set covers incretin therapeutics evidence.
Which model leads Health Optimization Bench?
Claude Fable 5 (Anthropic) holds the top current result on Health Optimization Bench at 83.8, per healthoptimizationbench.com, as of 2026-08.
Where do the Health Optimization Bench numbers come from?
From healthoptimizationbench.com (independently run). Cross-family authoring with blind three-family panel grading; the authoring family never grades its own task, and 95 percent bootstrap confidence intervals accompany every score on the site.
The rest of the field is on the index, and how sources qualify is on the methodology page. Model names in the table link to cross-benchmark pages.