Clinical Benchmarks

Gemma 4 E4B

Released 2 Apr 2026128K context8BopenCompare with other models

Clinical Benchmarks Index
5.2rank 139 of 148; 9.0 × 0.577 = 5.2, from 1 of 10 boards
Boards
1 of 13
Results
1
Latest measurement
Apr 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Clinical reasoning and knowledge

  1. MedXpertQA (MM)
    Google model card; MedXPertQA MM row; vendor-reported; protocol differs from other source families
    28.7
    Rank 21 of 22 here, 5 models on the boardLeader GPT-5.6 Sol 81.5
    Vendor-reported
    Measured Apr 2026

Sources

Open a line for the quote and page.

  1. 21MedXpertQA (MM) Google model card; MedXPertQA MM row; vendor-reported; protocol differs from ot… 28.7
    Printed as 28.7%Vendor-reported, measured Apr 2026Configuration: Google model card; MedXPertQA MM row; vendor-reported; protocol differs from other source families
    Gemma 4 model card model card, Google, 2 Apr 2026. Evaluation Results, MedXPertQA MM row, Gemma 4 E4B column.
    | Gemma 4 31B | Gemma 4 26B A4B | Gemma 4 12B Unified | Gemma 4 E4B | Gemma 4 E2B | Gemma 3 27B (no think) MedXPertQA MM | 61.3% | 58.1% | 48.7% | 28.7% | 23.5% | -
    Every result from this document

Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Flash Preview (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 12B, Gemma 4 26B A4B, Gemma 4 31B, Gemma 4 E2B, MedGemma 27B Text, MedGemma 4B