Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

Qwen: Qwen3.5-27B vs Qwen: Qwen3.7 Max

Last updated at: 2026-05-22

Metric Qwen3.5-27B Qwen3.5-27B medium Release: 2026-02-24 Qwen3.7 Max Qwen3.7 Max none Release: 2026-05-22
Score 7.9 7.9
Rank #26 #27
Reliability 10.0 10.0
Consistency 8.9 10.0
Tests Correct
Attempt pass rate 73.3% 70.0%
Flaky tests 3 0
Total Runs 60 60
Cost per result 4.664 0.719
Total Cost $0.607 $0.101
Input Price $0.195 / 1M $2.500 / 1M
Output Price $1.560 / 1M $7.500 / 1M
Output Tokens 2,572 1,988
Reasoning Tokens 312,011 0
Response Time (avg) 60.85s 1.30s
Response Time (max) 177.36s 3.92s
Response Time (total) 1216.93s 25.95s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 8.7 7.9 91.7% 1 19.75s 569 31,505
Qwen3.7 Max 6.5 10.0 50.0% 0 1.08s 242 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 7.0 9.8 50.0% 0 123.86s 416 64,993
Qwen3.7 Max 6.8 10.0 50.0% 0 1.39s 576 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 10.0 10.0 100.0% 0 163.96s 483 9,991
Qwen3.7 Max 3.0 10.0 0.0% 0 2.17s 171 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 10.0 10.0 100.0% 0 30.26s 270 16,150
Qwen3.7 Max 10.0 10.0 100.0% 0 1.35s 243 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 5.3 10.0 33.3% 0 79.53s 43 52,368
Qwen3.7 Max 7.7 10.0 66.7% 0 975ms 15 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 6.1 3.1 66.7% 1 101.41s 70 23,147
Qwen3.7 Max 10.0 10.0 100.0% 0 1.04s 120 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 10.0 10.0 100.0% 0 19.66s 97 11,638
Qwen3.7 Max 10.0 10.0 100.0% 0 943ms 72 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 8.2 7.7 77.8% 1 64.61s 245 77,213
Qwen3.7 Max 10.0 10.0 100.0% 0 1.13s 314 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 10.0 10.0 100.0% 0 7.45s 348 1,323
Qwen3.7 Max 10.0 10.0 100.0% 0 3.92s 222 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Qwen3.5-27B 3.0 10.0 0.0% 0 85.11s 31 23,683
Qwen3.7 Max 3.0 10.0 0.0% 0 856ms 13 0

Quick Compare

Switch Comparison Pair