Clinical Benchmarks

Gemini 2.5 Flash

Released 17 Jun 2025$0.3 input, $2.5 output per million tokens1M contextproprietaryAlso written as gemini-2.5-flash-preview-09-2025, Gemini 2.5 Flash (7/17) (Nonthinking), Gemini 2.5 Flash (7/17) (Thinking), google/gemini-2.5-flash-preview-09-2025, gemini-2.5-flash-preview-9-25, Gemini 2.5 Flash Preview (9/25) (Nonthinking), gemini-2.5-flash-7-17, google/gemini-2.5-flash, Gemini 2.5 Flash Preview (9/25) (Thinking), google/gemini-2.5-flash-thinking, gemini-2.5-flash-thinking, google/gemini-2.5-flash-preview-09-2025-thinking, gemini-2.5-flash-preview-09-2025-thinkingCompare with other models

Clinical Benchmarks Index
52.1rank 61 of 148; 90.3 × 0.577 = 52.1, from 1 of 10 boards
Boards
1 of 13
Results
4
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID google/gemini-2.5-flash-thinking; temperature=1; max_output_tokens=30000
    82.98
    Rank 45 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedScribe (Vals AI)
    model ID google/gemini-2.5-flash; temperature=1; max_output_tokens=30000
    82.87
    Rank 47 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  3. MedScribe (Vals AI)
    model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_output_tokens=30000
    78.50
    Rank 61 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  4. MedScribe (Vals AI)
    model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tokens=30000
    77.95
    Rank 64 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 45MedScribe (Vals AI) model ID google/gemini-2.5-flash-thinking; temperature=1; max_output_tokens=300… 82.98
    Printed as 82.98%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 45 of 105 (Gemini 2.5 Flash (7/17) (Thinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["google/gemini-2.5-flash-thinking"].
    45 | Gemini 2.5 Flash (7/17) (Thinking) | 82.98%±1.91 | $0.3/$2.5 | 22.53s
    Every result from this document
  2. 47MedScribe (Vals AI) model ID google/gemini-2.5-flash; temperature=1; max_output_tokens=30000 82.87
    Printed as 82.87%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 47 of 105 (Gemini 2.5 Flash (7/17) (Nonthinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["google/gemini-2.5-flash"].
    47 | Gemini 2.5 Flash (7/17) (Nonthinking) | 82.87%±1.91 | $0.3/$2.5 | 22.79s
    Every result from this document
  3. 61MedScribe (Vals AI) model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_o… 78.50
    Printed as 78.50%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash-preview-09-2025-thinking; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 61 of 105 (Gemini 2.5 Flash Preview (9/25) (Thinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["google/gemini-2.5-flash-preview-09-2025-thinking"].
    61 | Gemini 2.5 Flash Preview (9/25) (Thinking) | 78.50%±1.99 | $0.3/$2.5 | 31.39s
    Every result from this document
  4. 64MedScribe (Vals AI) model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tok… 77.95
    Printed as 77.95%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-2.5-flash-preview-09-2025; temperature=1; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 64 of 105 (Gemini 2.5 Flash Preview (9/25) (Nonthinking)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["google/gemini-2.5-flash-preview-09-2025"].
    64 | Gemini 2.5 Flash Preview (9/25) (Nonthinking) | 77.95%±1.92 | $0.3/$2.5 | 22.89s
    Every result from this document

Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Flash Preview (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 12B, Gemma 4 26B A4B, Gemma 4 31B, Gemma 4 E2B, Gemma 4 E4B, MedGemma 27B Text, MedGemma 4B