Gemma 4 31B
Released 2 Apr 2026256K context30.7BopenCompare with other models
- Clinical Benchmarks Index
- 37.6rank 103 of 148; 65.2 × 0.577 = 37.6, from 1 of 10 boards
- Boards
- 1 of 13
- Results
- 1
- Latest measurement
- Apr 2026
Results
Each line is placed on its own board. Dark tick: the board leader.
Clinical reasoning and knowledge
- MedXpertQA (MM)Google model card; MedXPertQA MM row; vendor-reported; protocol differs from other source families61.3Rank 17 of 22 here, 5 models on the boardLeader GPT-5.6 Sol 81.5
Sources
Open a line for the quote and page.
17MedXpertQA (MM) Google model card; MedXPertQA MM row; vendor-reported; protocol differs from ot… 61.3
Printed as 61.3%Vendor-reported, measured Apr 2026Configuration: Google model card; MedXPertQA MM row; vendor-reported; protocol differs from other source familiesGemma 4 model card model card, Google, 2 Apr 2026. Evaluation Results, MedXPertQA MM row, Gemma 4 31B column.| Gemma 4 31B | Gemma 4 26B A4B | Gemma 4 12B Unified | Gemma 4 E4B | Gemma 4 E2B | Gemma 3 27B (no think) MedXPertQA MM | 61.3% | 58.1% | 48.7% | 28.7% | 23.5% | -
Every result from this document
Other Google models: Gemini 2.0 Flash, Gemini 2.5 Flash, Gemini 2.5 Flash (7/17), Gemini 2.5 Flash Lite, Gemini 2.5 Flash Lite (9/25), Gemini 2.5 Flash Preview (9/25), Gemini 2.5 Pro, Gemini 3.1 Flash Lite Preview, Gemini 3.1 Pro, Gemini 3.5 Flash, Gemini 3.5 Flash Lite, Gemini 3.6 Flash, Gemini 3.7 Flash, Gemini 3.8 Flash, Gemini 3 Flash, Gemini 3 Pro, Gemini 3 Pro (11/25), Gemma 3 12B, Gemma 3 27B, Gemma 4 12B, Gemma 4 26B A4B, Gemma 4 E2B, Gemma 4 E4B, MedGemma 27B Text, MedGemma 4B