Clinical Benchmarks

OpenAI logoGPT-5.6 Luna: healthcare benchmark results

OpenAI · 1 board · updated August 16, 2026

The index currently holds one result for GPT-5.6 Luna: 55.8 on HealthBench (7 of 8). Scores below sit on different scales and come from different graders, so read each against its own benchmark, never against the others.

Results by benchmark

benchmarkscorepositionas of
HealthBench
via OpenAI Deployment Safety Hub (GPT-5.6 system card + August 2026 updates); benchlm.ai and llm-stats.com mirror
55.87 of 82026-06

Position counts against the source's full board, including rows this index does not mirror. Config caveats, where a source noted any, are on each benchmark's page.

Which healthcare benchmarks is GPT-5.6 Luna scored on?

As of August 16, 2026, GPT-5.6 Luna holds current results on 1 tracked benchmark: HealthBench.

How does GPT-5.6 Luna rank on them?

GPT-5.6 Luna stands at 55.8 on HealthBench (7 of 8).

The benchmarks themselves are described on their pages, linked in the table above, and the whole field is on the index.