Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Qwen: Qwen3.5-122B-A10B vs Qwen: Qwen3.5 Plus 2026-04-20

Last updated at: 2026-05-29

Metric Qwen3.5-122B-A10B Qwen3.5-122B-A10B none Release: 2026-02-24 Qwen3.5 Plus 2026-04-20 Qwen3.5 Plus 2026-04-20 none Release: 2026-04-20
Score 5.4 5.8
Rank #130 #109
Reliability 10.0 10.0
Consistency 9.5 8.5
Tests Correct
Attempt pass rate 33.3% 43.3%
Flaky tests 1 4
Total Runs 60 60
Cost per result 0.380 0.582
Total Cost $0.023 $0.041
Input Price $0.260 / 1M $0.300 / 1M
Output Price $2.080 / 1M $1.800 / 1M
Output Tokens 3,374 11,139
Reasoning Tokens 0 0
Response Time (avg) 3.38s 4.57s
Response Time (max) 46.00s 33.34s
Response Time (total) 67.55s 91.37s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 4.8 10.0 25.0% 0 1.59s 312 0
Qwen3.5 Plus 2026-04-20 4.8 10.0 25.0% 0 1.88s 557 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 4.0 5.5 33.3% 1 2.14s 684 0
Qwen3.5 Plus 2026-04-20 4.4 6.7 16.7% 1 2.08s 474 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 3.0 10.0 0.0% 0 46.00s 1,137 0
Qwen3.5 Plus 2026-04-20 2.8 1.6 33.3% 1 13.32s 2,275 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 1.01s 243 0
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 2.82s 243 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 5.3 10.0 33.3% 0 465ms 15 0
Qwen3.5 Plus 2026-04-20 5.3 10.0 33.3% 0 4.43s 18 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 5.0 10.0 0.0% 0 1.12s 66 0
Qwen3.5 Plus 2026-04-20 4.8 10.0 0.0% 0 1.41s 119 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 6.3 10.0 50.0% 0 513ms 69 0
Qwen3.5 Plus 2026-04-20 6.2 5.8 66.7% 1 1.17s 68 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 3.8 10.0 0.0% 0 1.00s 575 0
Qwen3.5 Plus 2026-04-20 6.7 7.9 55.6% 1 1.97s 583 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 10.0 10.0 100.0% 0 2.04s 264 0
Qwen3.5 Plus 2026-04-20 10.0 10.0 100.0% 0 4.42s 297 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-122B-A10B 3.0 10.0 0.0% 0 295ms 9 0
Qwen3.5 Plus 2026-04-20 3.0 10.0 0.0% 0 33.34s 6,505 0

Quick Compare

Switch Comparison Pair