Clinical Benchmarks

Gemini 3.7 Flash

Released 13 Aug 2026$0.75 input, $3.75 output per million tokens1M contextproprietaryAlso written as gemini-3.7-flashCompare with other models

Clinical Benchmarks Index
68.6rank 24 of 148; 84.1 × 0.816 = 68.6, from 2 of 10 boards
Boards
2 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID google/gemini-3.7-flash; temperature=1; max_output_tokens=30000; reasoning_effort=high
    83.94
    Rank 37 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID google/gemini-3.7-flash; reasoning_effort=high; temperature=1; max_output_tokens=30000
    53.39
    Rank 8 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 37MedScribe (Vals AI) model ID google/gemini-3.7-flash; temperature=1; max_output_tokens=30000; reaso… 83.94
    Printed as 83.94%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-3.7-flash; temperature=1; max_output_tokens=30000; reasoning_effort=high
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 37 of 105 (Gemini 3.7 Flash), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["google/gemini-3.7-flash"].
    37 | Gemini 3.7 Flash | 83.94%±2.00 | $1.5/$7.5 | 16.82s
    Every result from this document
  2. 8MedCode (Vals AI) model ID google/gemini-3.7-flash; reasoning_effort=high; temperature=1; max_out… 53.39
    Printed as 53.39%Official leaderboard, measured Sep 2026Configuration: model ID google/gemini-3.7-flash; reasoning_effort=high; temperature=1; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 8 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["google/gemini-3.7-flash"].
    8 | Gemini 3.7 Flash | 53.39%±2.12 | $1.5/$7.5 | 9.16s
    Every result from this document

Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Flash Preview (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 12B, Gemma 4 26B A4B, Gemma 4 31B, Gemma 4 E2B, Gemma 4 E4B, MedGemma 27B Text, MedGemma 4B