o3
Released 16 Apr 2025$2 input, $8 output per million tokens200K contextproprietaryAlso written as o3-2025-04-16, OpenAI o3, openai-o3, openai/o3-2025-04-16Compare with other models
- Clinical Benchmarks Index
- 59.6rank 45 of 148; 73.0 × 0.816 = 59.6, from 2 of 10 boards
- Boards
- 2 of 13
- Results
- 2
- Latest measurement
- Sep 2026
Results
Each line is placed on its own board. Dark tick: the board leader.
Documentation and coding
- MedScribe (Vals AI)model ID openai/o3-2025-04-16; max_output_tokens=30000; reasoning_effort=high76.65Rank 70 of 105Leader Claude Opus 5.5 91.43
- MedCode (Vals AI)model ID openai/o3-2025-04-16; reasoning_effort=high; max_output_tokens=3000047.29Rank 31 of 103Leader Claude Opus 5 63.57
Sources
Open a line for the quote and page.
70MedScribe (Vals AI) model ID openai/o3-2025-04-16; max_output_tokens=30000; reasoning_effort=high 76.65
Printed as 76.65%Official leaderboard, measured Sep 2026Configuration: model ID openai/o3-2025-04-16; max_output_tokens=30000; reasoning_effort=highVals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 70 of 105 (o3), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["openai/o3-2025-04-16"].70 | o3 | 76.65%±1.87 | $2/$8 | 46.39s
Every result from this document31MedCode (Vals AI) model ID openai/o3-2025-04-16; reasoning_effort=high; max_output_tokens=30000 47.29
Printed as 47.29%Official leaderboard, measured Sep 2026Configuration: model ID openai/o3-2025-04-16; reasoning_effort=high; max_output_tokens=30000Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 31 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["openai/o3-2025-04-16"].31 | o3 | 47.29%±2.16 | $2/$8 | 17.68s
Every result from this document
Other OpenAI models: GPT-4.1, GPT-4.1 mini, GPT-4o, GPT-5, GPT-5.1, GPT-5.2, GPT-5.3-Codex, GPT-5.4, GPT-5.4 mini, GPT-5.4 nano, GPT-5.5, GPT-5.5 Instant, GPT-5.6 Luna, GPT-5.6 Sol, GPT-5.6 Terra, GPT-5 mini, GPT-5 nano, GPT-6.1 Sol, GPT-6 Astra, GPT-6 Luna, GPT-6 Sol, GPT OSS 120B, GPT OSS 20B, o4-mini