Clinical Benchmarks

Qwen 3.8 27B

256K context27BopenAlso written as alibaba/qwen3.8-27b, qwen3.8-27bCompare with other models

Clinical Benchmarks Index
45.6rank 76 of 148; 55.9 × 0.816 = 45.6, from 2 of 10 boards
Boards
2 of 13
Results
2
Latest measurement
Sep 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

Documentation and coding

  1. MedScribe (Vals AI)
    model ID alibaba/qwen3.8-27b; temperature=1; top_p=0.95; max_output_tokens=30000; reasoning_effort=xhigh
    83.85
    Rank 38 of 105Leader Claude Opus 5.5 91.43
    Official leaderboard
    Measured Sep 2026
  2. MedCode (Vals AI)
    model ID alibaba/qwen3.8-27b; reasoning_effort=xhigh; temperature=1; top_p=0.95; max_output_tokens=30000
    28.70
    Rank 95 of 103Leader Claude Opus 5 63.57
    Official leaderboard
    Measured Sep 2026

Sources

Open a line for the quote and page.

  1. 38MedScribe (Vals AI) model ID alibaba/qwen3.8-27b; temperature=1; top_p=0.95; max_output_tokens=3000… 83.85
    Printed as 83.85%Official leaderboard, measured Sep 2026Configuration: model ID alibaba/qwen3.8-27b; temperature=1; top_p=0.95; max_output_tokens=30000; reasoning_effort=xhigh
    Vals AI MedScribe leaderboard official leaderboard, Vals AI, 3 Sep 2026. Vals AI MedScribe leaderboard, View: All Models, Task: Overall, row 38 of 105 (Qwen 3.8 27B), Accuracy column; Updated 9/29/2026. Rendered BenchmarkView table; configuration from embedded astro-island BenchmarkView props, benchmarkView.default.tasks.overall["alibaba/qwen3.8-27b"].
    38 | Qwen 3.8 27B | 83.85%±1.98 | $0.5/$3 | 89.70s
    Every result from this document
  2. 95MedCode (Vals AI) model ID alibaba/qwen3.8-27b; reasoning_effort=xhigh; temperature=1; top_p=0.95… 28.70
    Printed as 28.70%Official leaderboard, measured Sep 2026Configuration: model ID alibaba/qwen3.8-27b; reasoning_effort=xhigh; temperature=1; top_p=0.95; max_output_tokens=30000
    Vals AI MedCode leaderboard official leaderboard, Vals AI, 29 Sep 2026. MedCode leaderboard, Overall task, All Models expanded, rank 95 of 103; Accuracy column; board Updated 9/29/2026. Rendered table row; configuration from embedded BenchmarkView props benchmarkView.default.tasks.overall["alibaba/qwen3.8-27b"].
    95 | Qwen 3.8 27B | 28.70%±1.97 | $0.5/$3 | 89.38s
    Every result from this document

Other Alibaba models: Lingshu-32B, Lingshu-7B, Qwen3-14B, Qwen3-235B-A22B-Instruct-2507, Qwen3-32B, Qwen3-4B, Qwen 3.5, Qwen3.5-27B, Qwen3.5-35B-A3B, Qwen3.5 397B A17B, Qwen3.5-9B, Qwen 3.5 Flash, Qwen3.5-Plus, Qwen3.6-Max, Qwen3.6 Plus, Qwen 3.7 Max, Qwen3.7 Plus, Qwen3.8 Max, Qwen 3 Max Thinking, Qwen3-VL-235B-A22B, Qwen 3 VL Plus