GPT-5.4 (low reasoning): healthcare benchmark results
OpenAI · 1 board · updated August 16, 2026
The index currently holds one result for GPT-5.4 (low reasoning): 0.58 on EHR-Complex (5 of 12). Scores below sit on different scales and come from different graders, so read each against its own benchmark, never against the others.
Results by benchmark
| benchmark | score | position | as of |
|---|---|---|---|
| EHR-Complex via EHR-Complex paper | 0.58 | 5 of 12 | 2026-06 |
Position counts against the source's full board, including rows this index does not mirror. Config caveats, where a source noted any, are on each benchmark's page.
Which healthcare benchmarks is GPT-5.4 (low reasoning) scored on?
As of August 16, 2026, GPT-5.4 (low reasoning) holds current results on 1 tracked benchmark: EHR-Complex.
How does GPT-5.4 (low reasoning) rank on them?
GPT-5.4 (low reasoning) stands at 0.58 on EHR-Complex (5 of 12).
The benchmarks themselves are described on their pages, linked in the table above, and the whole field is on the index.