Clinical Benchmarks

GPT-5.4 nano

Released 17 Mar 2026$0.2 input, $1.25 output per million tokens400K contextproprietaryAlso written as gpt-5.4-nano-2026-03-17, GPT 5.4 Nano, openai/gpt-5.4-nano-2026-03-17Compare with other models

Clinical Benchmarks Index
53.9rank 56 of 148; 66.1 × 0.816 = 53.9, from 2 of 10 boards
Boards
2 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID openai/gpt-5.4-nano-2026-03-17; max_output_tokens=30000; reasoning_effort=high
    77.09
    Rank 68 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID openai/gpt-5.4-nano-2026-03-17; reasoning_effort=high; max_output_tokens=30000
    41.03
    Rank 56 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 68MedScribe (Vals AI) model ID openai/gpt-5.4-nano-2026-03-17; max_output_tokens=30000; reasoning_eff… 77.09
    Printed as 77.09%Official leaderboard, measured Sep 2026Configuration: model ID openai/gpt-5.4-nano-2026-03-17; max_output_tokens=30000; reasoning_effort=high
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 68 of 105 (GPT 5.4 Nano), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["openai/gpt-5.4-nano-2026-03-17"].
    68 | GPT 5.4 Nano | 77.09%±1.89 | $0.2/$1.25 | 20.53s
    Every result from this document
  2. 56MedCode (Vals AI) model ID openai/gpt-5.4-nano-2026-03-17; reasoning_effort=high; max_output_toke… 41.03
    Printed as 41.03%Official leaderboard, measured Sep 2026Configuration: model ID openai/gpt-5.4-nano-2026-03-17; reasoning_effort=high; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 56 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["openai/gpt-5.4-nano-2026-03-17"].
    56 | GPT 5.4 Nano | 41.03%±2.26 | $0.2/$1.25 | 8.04s
    Every result from this document

Other OpenAI models: GPT-4.1, GPT-4.1 mini, GPT-4o, GPT-5, GPT-5.1, GPT-5.2, GPT-5.3-Codex, GPT-5.4, GPT-5.4 mini, GPT-5.5, GPT-5.5 Instant, GPT-5.6 Luna, GPT-5.6 Sol, GPT-5.6 Terra, GPT-5 mini, GPT-5 nano, GPT-6.1 Sol, GPT-6 Astra, GPT-6 Luna, GPT-6 Sol, GPT OSS 120B, GPT OSS 20B, o3, o4-mini