Clinical Benchmarks

Qwen3.6-Max

$1.3 input, $7.8 output per million tokens256K contextproprietaryAlso written as qwen3.6-max, qwen3-6-max-previewCompare with other models

Boards
1 of 13
Results
4
Latest measurement
May 2026

Results

Each line is placed on its own board. Dark tick: the board leader.

EHR and workflow agents

  1. CHI-Bench
    Listed as hermes + qwen-3.6-max
    All Domains pass@1; hermes harness; Qwen3.6-Max preview endpoint
    16.4
    Rank 19 of 43 here, 45 models on the boardLeader erius + claude-opus-5 54.7
    Official leaderboard
    Measured May 2026
  2. CHI-Bench
    Listed as openai-agents + qwen-3.6-max
    All Domains pass@1; openai-agents harness; Qwen3.6-Max preview endpoint
    15.6
    Rank 21 of 43 here, 45 models on the boardLeader erius + claude-opus-5 54.7
    Official leaderboard
    Measured May 2026
  3. CHI-Bench
    Listed as deepagents + qwen-3.6-max
    All Domains pass@1; deepagents harness; Qwen3.6-Max preview endpoint
    9.3
    Rank 33 of 43 here, 45 models on the boardLeader erius + claude-opus-5 54.7
    Official leaderboard
    Measured May 2026
  4. CHI-Bench
    Listed as openclaw + qwen-3.6-max
    All Domains pass@1; openclaw harness; Qwen3.6-Max preview endpoint
    4.9
    Rank 38 of 43 here, 45 models on the boardLeader erius + claude-opus-5 54.7
    Official leaderboard
    Measured May 2026

Sources

Open a line for the quote and page.

  1. 19CHI-Bench All Domains pass@1; hermes harness; Qwen3.6-Max preview endpoint 16.4
    Printed as 16.4%Official leaderboard, measured May 2026Configuration: All Domains pass@1; hermes harness; Qwen3.6-Max preview endpoint
    CHI-Bench leaderboard (actAVA) official leaderboard, actAVA, 12 Aug 2026. CHI-Bench v1.0.0, All Domains, rank 19, Agent/Model and Accuracy columns; PA/UM/CM follow; board last updated 2026-08-12; board Date column 2026-05-01; accessed 2026-09-30. Linked submission provenance: run 2026-05-03T07:32:36Z to 2026-05-04T06:18:19Z; submitted_at 2026-05-04T06:18:19Z.
    19 | hermes | qwen-3.6-max | Open-source | 16.4% | 9.3% | 26.7% | 13.3% | 2026-05-01
    Every result from this document
  2. 21CHI-Bench All Domains pass@1; openai-agents harness; Qwen3.6-Max preview endpoint 15.6
    Printed as 15.6%Official leaderboard, measured May 2026Configuration: All Domains pass@1; openai-agents harness; Qwen3.6-Max preview endpoint
    CHI-Bench leaderboard (actAVA) official leaderboard, actAVA, 12 Aug 2026. CHI-Bench v1.0.0, All Domains, rank 21, Agent/Model and Accuracy columns; PA/UM/CM follow; board last updated 2026-08-12; board Date column 2026-05-01; accessed 2026-09-30. Linked submission provenance: run 2026-05-03T09:19:32Z to 2026-05-04T07:25:50Z; submitted_at 2026-05-04T07:25:50Z.
    21 | openai-agents | qwen-3.6-max | Open-source | 15.6% | 16.0% | 26.7% | 4.0% | 2026-05-01
    Every result from this document
  3. 33CHI-Bench All Domains pass@1; deepagents harness; Qwen3.6-Max preview endpoint 9.3
    Printed as 9.3%Official leaderboard, measured May 2026Configuration: All Domains pass@1; deepagents harness; Qwen3.6-Max preview endpoint
    CHI-Bench leaderboard (actAVA) official leaderboard, actAVA, 12 Aug 2026. CHI-Bench v1.0.0, All Domains, rank 33, Agent/Model and Accuracy columns; PA/UM/CM follow; board last updated 2026-08-12; board Date column 2026-05-01; accessed 2026-09-30. Linked submission provenance: run 2026-05-03T22:18:28Z to 2026-05-04T13:19:53Z; submitted_at 2026-05-04T13:19:53Z.
    33 | deepagents | qwen-3.6-max | Open-source | 9.3% | 12.0% | 10.7% | 5.3% | 2026-05-01
    Every result from this document
  4. 38CHI-Bench All Domains pass@1; openclaw harness; Qwen3.6-Max preview endpoint 4.9
    Printed as 4.9%Official leaderboard, measured May 2026Configuration: All Domains pass@1; openclaw harness; Qwen3.6-Max preview endpoint
    CHI-Bench leaderboard (actAVA) official leaderboard, actAVA, 12 Aug 2026. CHI-Bench v1.0.0, All Domains, rank 39, Agent/Model and Accuracy columns; PA/UM/CM follow; board last updated 2026-08-12; board Date column 2026-05-01; accessed 2026-09-30. Linked submission provenance: run 2026-05-02T09:26:43Z to 2026-05-04T03:15:50Z; submitted_at 2026-05-04T03:15:50Z.
    39 | openclaw | qwen-3.6-max | Open-source | 4.9% | 10.7% | 4.0% | 0.0% | 2026-05-01
    Every result from this document

Other Alibaba models: Lingshu-32B, Lingshu-7B, Qwen3-14B, Qwen3-235B-A22B-Instruct-2507, Qwen3-32B, Qwen3-4B, Qwen 3.5, Qwen3.5-27B, Qwen3.5-35B-A3B, Qwen3.5 397B A17B, Qwen3.5-9B, Qwen 3.5 Flash, Qwen3.5-Plus, Qwen3.6 Plus, Qwen 3.7 Max, Qwen3.7 Plus, Qwen 3.8 27B, Qwen3.8 Max, Qwen 3 Max Thinking, Qwen3-VL-235B-A22B, Qwen 3 VL Plus