Clinical Benchmarks

Microsoft/OpenAI logoCopilot (GPT 5.5): healthcare benchmark results

Microsoft/OpenAI · 1 board · updated August 16, 2026

The index currently holds one result for Copilot (GPT 5.5): 35% on HealthAgentBench (5 of 12). Scores below sit on different scales and come from different graders, so read each against its own benchmark, never against the others.

Results by benchmark

benchmarkscorepositionas of
HealthAgentBench
via HealthAgentBench leaderboard (Microsoft GitHub Pages)
35%5 of 122026-07

Position counts against the source's full board, including rows this index does not mirror. Config caveats, where a source noted any, are on each benchmark's page.

Which healthcare benchmarks is Copilot (GPT 5.5) scored on?

As of August 16, 2026, Copilot (GPT 5.5) holds current results on 1 tracked benchmark: HealthAgentBench.

How does Copilot (GPT 5.5) rank on them?

Copilot (GPT 5.5) stands at 35% on HealthAgentBench (5 of 12).

The benchmarks themselves are described on their pages, linked in the table above, and the whole field is on the index.