Clinical Benchmarks

MiniMax-M2.7

Released 18 Mar 2026$0.3 input, $1.2 output per million tokensopenAlso written as minimax/MiniMax-M2.7Compare with other models

Clinical Benchmarks Index
41.9rank 94 of 148; 41.9 × 1 = 41.9, from 3 of 10 boards
Boards
3 of 13
Results
3
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=30000
    79.87
    Rank 56 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=30000
    34.44
    Rank 76 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

EHR and workflow agents

  1. PhysicianBench
    Pass@1 over 3 runs; Pass^3 1.0; shared FHIR tool harness; up to 100 turns; high reasoning when supported
    8.7
    Rank 19 of 21 here, 12 models on the boardLeader Claude Opus 5.5 (max) 68.4
    Official leaderboard
    Measured May 2026

Sources

Open a line for the quote and page.

  1. 56MedScribe (Vals AI) model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=300… 79.87
    Printed as 79.87%Official leaderboard, measured Sep 2026Configuration: model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 56 of 105 (MiniMax-M2.7), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["minimax/MiniMax-M2.7"].
    56 | MiniMax-M2.7 | 79.87%±1.86 | $0.3/$1.2 | 27.25s
    Every result from this document
  2. 76MedCode (Vals AI) model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=300… 34.44
    Printed as 34.44%Official leaderboard, measured Sep 2026Configuration: model ID minimax/MiniMax-M2.7; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 76 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["minimax/MiniMax-M2.7"].
    76 | MiniMax-M2.7 | 34.44%±1.99 | $0.3/$1.2 | 30.25s
    Every result from this document
  3. 19PhysicianBench Pass@1 over 3 runs; Pass^3 1.0; shared FHIR tool harness; up to 100 turns; high… 8.7
    Printed as 8.7 ± 1.2Official leaderboard, measured May 2026Configuration: Pass@1 over 3 runs; Pass^3 1.0; shared FHIR tool harness; up to 100 turns; high reasoning when supported
    PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments (arXiv 2605.02240v1 PDF) paper, Stanford University (HealthRex; Liu, Chen et al.), 4 May 2026. arXiv HTML 2605.02240v1, Section 5.2, Table 2; MiniMax M2.7 row, Pass@1 column (PDF p. 8).
    Model | Pass@1 | Pass@3 | Pass^3 | #Turns MiniMax M2.7 | 8.7 ± 1.2 | 15.9 | 1.0 | 29.7
    Every result from this document

Other MiniMax models: MiniMax-M2.1, MiniMax M3