Clinical Benchmarks

DeepSeek V4 Flash

Released 24 Apr 20261M context284B A13BopenAlso written as deepseek/deepseek-v4-flash-0731, deepseek-v4-flash-0731, DeepSeek V4 Flash 0731Compare with other models

Clinical Benchmarks Index
50.4rank 65 of 148; 87.3 × 0.577 = 50.4, from 1 of 10 boards
Boards
1 of 13
Results
1
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID deepseek/deepseek-v4-flash-0731; max_output_tokens=30000; reasoning_effort=high
    80.36
    Rank 54 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 54MedScribe (Vals AI) model ID deepseek/deepseek-v4-flash-0731; max_output_tokens=30000; reasoning_ef… 80.36
    Printed as 80.36%Official leaderboard, measured Sep 2026Configuration: model ID deepseek/deepseek-v4-flash-0731; max_output_tokens=30000; reasoning_effort=high
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 54 of 105 (DeepSeek V4 Flash 0731), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["deepseek/deepseek-v4-flash-0731"].
    54 | DeepSeek V4 Flash 0731 | 80.36%±1.97 | $0.44/$1.32 | 77.55s
    Every result from this document

Other DeepSeek models: DeepSeek R1, DeepSeek-V3.1, DeepSeek-V3.2-Exp, DeepSeek V4.1 Flash, DeepSeek V4 Flash 0731, DeepSeek V4 Pro