Qwen3.5-Plus
$0.12 input, $0.69 output per million tokens1M contextproprietaryCompare with other models
- Clinical Benchmarks Index
- 47.0rank 71 of 148; 81.4 × 0.577 = 47.0, from 1 of 10 boards
- Boards
- 1 of 13
- Results
- 76 on secondary measures
- Latest measurement
- Aug 2026
Results
Each line is placed on its own board. Dark tick: the board leader.
Safety
- MedPICzero-shot; independent questions; exact option-set match72.4Rank 3 of 28Leader Gemini-3.1-Pro 80.7
Secondary measures
Safety
- 81.7Rank 3 of 28Leader Gemini-3.1-Pro 87.7
- 57.9Rank 3 of 28Leader Gemini-3.1-Pro 69.9
- 32.6Rank 3 of 28Leader Gemini-3.1-Pro 48.3
- 55.9Rank 3 of 28Leader Gemini-3.1-Pro 77.9
- 66.2Rank 7 of 28Leader MedGemma-27B-Text 76.1
- 23.8Descriptive measure, not ranked
Sources
Open a line for the quote and page.
3MedPIC zero-shot; independent questions; exact option-set match 72.4
Printed as 72.4Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column Overall; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document3MedPIC, Guideline-following accuracy (GF) zero-shot; independent questions; exact option-set match 81.7
Printed as 81.7Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column GF; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document3MedPIC, Counterfactual accuracy (CF) zero-shot; independent questions; exact option-set match 57.9
Printed as 57.9Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column CF; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document3MedPIC, Linked counterfactual pair accuracy zero-shot; independent questions; exact option-set match 32.6
Printed as 32.6Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column Pair; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document7MedPIC, Risk activation accuracy zero-shot; independent questions; exact option-set match 66.2
Printed as 66.2Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column Activation; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document3MedPIC, Risk deactivation accuracy zero-shot; independent questions; exact option-set match 55.9
Printed as 55.9Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column Deactivation; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document–MedPIC, Published GF-minus-CF accuracy gap zero-shot; independent questions; exact option-set match 23.8
Printed as 23.8Independent run, measured Aug 2026Configuration: zero-shot; independent questions; exact option-set matchEvaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning paper, Zhitian Hou, Yuhang Liu, Pengkai Wang, Zeyu Liu, Guanghao Zhu, Zheng Liu, Shuo Cai, Congkai Xie, Zhijie Sang, Kun Zeng, Hongxia Yang (The Hong Kong Polytechnic University; InfiX.ai; Sun Yat-sen University), 4 Aug 2026. arXiv:2608.03028v1, Section 4, Table 2 (#S4.T2), row Qwen3.5-Plus, column Δ_GF−CF; row order: Model | Overall | GF | CF | Δ_GF−CF | Activation | Deactivation | PairQwen3.5-Plus | 72.4 | 81.7 | 57.9 | 23.8 | 66.2 | 55.9 | 32.6
Every result from this document
Other Alibaba models: Lingshu-32B, Lingshu-7B, Qwen3-14B, Qwen3-235B-A22B-Instruct-2507, Qwen3-32B, Qwen3-4B, Qwen 3.5, Qwen3.5-27B, Qwen3.5-35B-A3B, Qwen3.5 397B A17B, Qwen3.5-9B, Qwen 3.5 Flash, Qwen3.6-Max, Qwen3.6 Plus, Qwen 3.7 Max, Qwen3.7 Plus, Qwen 3.8 27B, Qwen3.8 Max, Qwen 3 Max Thinking, Qwen3-VL-235B-A22B, Qwen 3 VL Plus