Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Qwen: Qwen3.5-35B-A3B vs Qwen: Qwen3.5 Plus 2026-04-20

Last updated at: 2026-05-22

Metric Qwen3.5-35B-A3B Qwen3.5-35B-A3B none Release: 2026-02-24 Qwen3.5 Plus 2026-04-20 Qwen3.5 Plus 2026-04-20 none Release: 2026-04-20
Score 5.8 5.8
Rank #102 #101
Reliability 10.0 9.9
Consistency 8.9 8.5
Tests Correct
Attempt pass rate 45.0% 43.3%
Flaky tests 3 4
Total Runs 60 60
Cost per result 0.224 0.583
Total Cost $0.016 $0.041
Input Price $0.139 / 1M $0.300 / 1M
Output Price $1.000 / 1M $1.800 / 1M
Output Tokens 4,334 11,174
Reasoning Tokens 0 0
Response Time (avg) 3.50s 4.58s
Response Time (max) 47.43s 33.34s
Response Time (total) 69.99s 91.55s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 3.4 7.9 16.7% 1 1.43s 574 0
Qwen3.5 Plus 2026-04-20 4.8 10.0 25.0% 0 1.88s 557 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 6.8 10.0 50.0% 0 1.72s 562 0
Qwen3.5 Plus 2026-04-20 4.4 6.7 16.7% 1 2.08s 474 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 3.0 10.0 0.0% 0 47.43s 1,833 0
Qwen3.5 Plus 2026-04-20 2.8 1.6 33.3% 1 13.32s 2,275 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 10.0 10.0 100.0% 0 1.16s 243 0
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 2.82s 243 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 7.7 10.0 66.7% 0 485ms 15 0
Qwen3.5 Plus 2026-04-20 5.3 10.0 33.3% 0 4.43s 18 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 6.5 3.4 66.7% 1 1.19s 114 0
Qwen3.5 Plus 2026-04-20 4.8 10.0 0.0% 0 1.41s 119 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 6.3 10.0 50.0% 0 809ms 63 0
Qwen3.5 Plus 2026-04-20 6.2 5.8 66.7% 1 1.17s 68 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 3.7 7.4 22.2% 1 1.34s 655 0
Qwen3.5 Plus 2026-04-20 6.7 7.9 55.6% 1 2.03s 618 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 10.0 10.0 100.0% 0 2.30s 264 0
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 4.42s 297 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 3.0 10.0 0.0% 0 493ms 11 0
Qwen3.5 Plus 2026-04-20 3.0 10.0 0.0% 0 33.34s 6,505 0

Quick Compare

Switch Comparison Pair