Clinical Benchmarks

OpenAI logoGPT-5.4 (low reasoning): healthcare benchmark results

OpenAI · 1 board · updated August 16, 2026

The index currently holds one result for GPT-5.4 (low reasoning): 0.58 on EHR-Complex (5 of 12). Scores below sit on different scales and come from different graders, so read each against its own benchmark, never against the others.

Results by benchmark

benchmarkscorepositionas of
EHR-Complex
via EHR-Complex paper
0.585 of 122026-06

Position counts against the source's full board, including rows this index does not mirror. Config caveats, where a source noted any, are on each benchmark's page.

Which healthcare benchmarks is GPT-5.4 (low reasoning) scored on?

As of August 16, 2026, GPT-5.4 (low reasoning) holds current results on 1 tracked benchmark: EHR-Complex.

How does GPT-5.4 (low reasoning) rank on them?

GPT-5.4 (low reasoning) stands at 0.58 on EHR-Complex (5 of 12).

The benchmarks themselves are described on their pages, linked in the table above, and the whole field is on the index.