Clinical Benchmarks

Gemma 4 12B

Released 3 Jun 2026256K context11.95BApache 2.0Also written as gemma-4-12b-it, Gemma 4 12B UnifiedCompare with other models

Clinical Benchmarks Index
25.0rank 120 of 148; 43.4 × 0.577 = 25.0, from 1 of 10 boards
Boards
1 of 13
Results
1
Latest measurement
Apr 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Clinical reasoning and knowledge

  1. MedXpertQA (MM)
    Google's Gemma 4 model card, Unified 12B; protocol not stated
    48.7
    Rank 19 of 22 here, 5 models on the boardLeader GPT-5.6 Sol 81.5
    Vendor-reported
    Measured Apr 2026

Sources

Open a line for the quote and page.

  1. 19MedXpertQA (MM) Google's Gemma 4 model card, Unified 12B; protocol not stated 48.7
    Printed as 48.7%Vendor-reported, measured Apr 2026Configuration: Google's Gemma 4 model card, Unified 12B; protocol not stated
    Gemma 4 model card model card, Google, 2 Apr 2026. Benchmark Results table, Vision section, row MedXPertQA MM, column Gemma 4 12B Unified; header row: | | Gemma 4 31B | Gemma 4 26B A4B | Gemma 4 12B Unified | Gemma 4 E4B | Gemma 4 E2B | Gemma 3 27B (no think) |
    | MedXPertQA MM | 61.3% | 58.1% | 48.7% | 28.7% | 23.5% | \- |
    Every result from this document

Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Flash Preview (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 26B A4B, Gemma 4 31B, Gemma 4 E2B, Gemma 4 E4B, MedGemma 27B Text, MedGemma 4B