Clinical Benchmarks

Grok 4.1 Fast (Reasoning)

Released 19 Nov 20252M contextproprietaryAlso written as grok/grok-4-1-fast-reasoning, grok-4-1-fast-reasoningCompare with other models

Clinical Benchmarks Index
42.7rank 88 of 148; 52.3 × 0.816 = 42.7, from 2 of 10 boards
Boards
2 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_tokens=30000
    78.73
    Rank 60 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_tokens=30000
    28.08
    Rank 97 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 60MedScribe (Vals AI) model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_to… 78.73
    Printed as 78.73%Official leaderboard, measured Sep 2026Configuration: model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 60 of 105 (Grok 4.1 Fast (Reasoning)), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["grok/grok-4-1-fast-reasoning"].
    60 | Grok 4.1 Fast (Reasoning) | 78.73%±1.87 | $0.2/$0.5 | 39.29s
    Every result from this document
  2. 97MedCode (Vals AI) model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_to… 28.08
    Printed as 28.08%Official leaderboard, measured Sep 2026Configuration: model ID grok/grok-4-1-fast-reasoning; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 97 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["grok/grok-4-1-fast-reasoning"].
    97 | Grok 4.1 Fast (Reasoning) | 28.08%±1.99 | $0.2/$0.5 | 46.50s
    Every result from this document

Other SpaceX AI models: Grok 4, Grok 4.1 Fast Non-Reasoning, Grok 4.20, Grok 4.3, Grok 4.5, Grok 4.6, Grok 4.7, Grok 4 Fast (Non-Reasoning), Grok 4 Fast (Reasoning)