Clinical Benchmarks

Gemini 2.5 Flash Preview (9/25)

Released 25 Sep 2025proprietaryAlso written as google/gemini-2.5-flash-preview-09-2025, Gemini 2.5 Flash Preview (9/25) (Nonthinking), Gemini 2.5 Flash Preview (9/25) (Thinking), google/gemini-2.5-flash-preview-09-2025-thinkingCompare with other models

Clinical Benchmarks Index
27.4rank 116 of 148; 47.5 × 0.577 = 27.4, from 1 of 10 boards
Boards
1 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedCode (Vals AI)
    model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tokens=30000
    40.54
    Rank 60 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_output_tokens=30000
    40.33
    Rank 63 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 60MedCode (Vals AI) model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tok… 40.54
    Printed as 40.54%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 60 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["google/gemini-2.5-flash-preview-09-2025"].
    60 | Gemini 2.5 Flash Preview (9/25) (Nonthinking) | 40.54%±1.93 | $0.3/$2.5 | 12.70s
    Every result from this document
  2. 63MedCode (Vals AI) model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_o… 40.33
    Printed as 40.33%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 63 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["google/gemini-2.5-flash-preview-09-2025-thinking"].
    63 | Gemini 2.5 Flash Preview (9/25) (Thinking) | 40.33%±1.92 | $0.3/$2.5 | 16.11s
    Every result from this document

Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 12B, Gemma 4 26B A4B, Gemma 4 31B, Gemma 4 E2B, Gemma 4 E4B, MedGemma 27B Text, MedGemma 4B