Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Qwen: Qwen3.5 Plus 2026-04-20 vs Qwen: Qwen3.6 Plus

Last updated at: 2026-05-22

Metric Qwen3.5 Plus 2026-04-20 Qwen3.5 Plus 2026-04-20 medium Release: 2026-04-20 Qwen3.6 Plus Qwen3.6 Plus medium Release: 2026-04-20
Score 7.6 7.8
Rank #42 #33
Reliability 9.6 10.0
Consistency 8.7 9.2
Tests Correct
Attempt pass rate 71.7% 68.3%
Flaky tests 3 2
Total Runs 60 60
Cost per result 2.789 0.630
Total Cost $0.363 $0.082
Input Price $0.300 / 1M $0.325 / 1M
Output Price $1.800 / 1M $1.950 / 1M
Output Tokens 2,245 1,822
Reasoning Tokens 150,235 124,938
Response Time (avg) 43.63s 26.78s
Response Time (max) 189.38s 201.68s
Response Time (total) 872.61s 508.74s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 10.84s 215 7,748
Qwen3.6 Plus 10.0 10.0 100.0% 0 9.90s 207 7,557
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 5.4 6.0 66.7% 1 137.55s 287 42,318
Qwen3.6 Plus 4.1 6.7 16.7% 1 201.68s 38 33,395
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 92.41s 483 17,490
Qwen3.6 Plus 10.0 10.0 100.0% 0 34.95s 452 13,073
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 38.32s 270 14,668
Qwen3.6 Plus 10.0 10.0 100.0% 0 14.95s 270 10,706
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 2.9 7.2 11.1% 1 53.10s 63 28,414
Qwen3.6 Plus 2.9 7.2 11.1% 1 29.59s 56 33,464
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 4.9 9.6 0.0% 0 25.30s 125 4,792
Qwen3.6 Plus 5.1 10.0 0.0% 0 27.05s 111 5,232
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 20.25s 103 7,689
Qwen3.6 Plus 10.0 10.0 100.0% 0 7.54s 102 5,552
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 8.2 7.2 88.9% 1 17.58s 324 9,786
Qwen3.6 Plus 10.0 10.0 100.0% 0 6.11s 298 6,868
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 14.72s 348 2,164
Qwen3.6 Plus 10.0 10.0 100.0% 0 5.87s 267 1,330
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5 Plus 2026-04-20 3.0 10.0 0.0% 0 92.57s 27 15,166
Qwen3.6 Plus 3.0 10.0 0.0% 0 47.51s 21 7,761

Quick Compare

Switch Comparison Pair