Clinical Benchmarks

OpenAI logoGPT-5.4 (high reasoning): healthcare benchmark results

OpenAI · 1 board · updated August 16, 2026

The index currently holds one result for GPT-5.4 (high reasoning): 0.65 on EHR-Complex (1 of 12). It tops EHR-Complex. Scores below sit on different scales and come from different graders, so read each against its own benchmark, never against the others.

Results by benchmark

benchmarkscorepositionas of
EHR-Complex
via EHR-Complex paper
0.651 of 122026-06

Position counts against the source's full board, including rows this index does not mirror. Config caveats, where a source noted any, are on each benchmark's page.

Which healthcare benchmarks is GPT-5.4 (high reasoning) scored on?

As of August 16, 2026, GPT-5.4 (high reasoning) holds current results on 1 tracked benchmark: EHR-Complex.

How does GPT-5.4 (high reasoning) rank on them?

GPT-5.4 (high reasoning) stands at 0.65 on EHR-Complex (1 of 12). It holds first place on EHR-Complex.

The benchmarks themselves are described on their pages, linked in the table above, and the whole field is on the index.