Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Qwen3.5-35B-A3B (medium) vs Qwen3.7 Flash

The average score is effectively tied at 6.2 vs 6.1. Qwen3.7 Flash has the lower benchmark cost at $0.019 vs $0.837. Qwen3.7 Flash is faster at 10.06s vs 112.47s, with pass rates of 66.7% vs 36.4%.

Last updated at: 2026-07-28

Rank
#133
Total Output Tokens
826,670
Response Time (avg)
112.47s
Total Cost
$0.837
Rank
#136
Total Output Tokens
90,507
Response Time (avg)
10.06s
Total Cost
$0.019
Recommended model Qwen3.7 Flash

It has the best score here (6.1), while costing about 45.6x less than Qwen3.5-35B-A3B (medium).

Detailed comparison

Metric Qwen3.5-35B-A3B Qwen3.5-35B-A3B medium Release: 2026-02-24 Qwen3.7 Flash Qwen3.7 Flash none Release: 2026-07-28
Score 6.2 6.1
Rank #133 #136
Reliability 10.0 10.0
Consistency 7.6 9.2
Tests Correct
Attempt pass rate 66.7% 36.4%
Flaky tests 6 2
Total Runs 66 66
Cost per result 9.130 0.262
Total Cost $0.837 $0.019
Input Price $0.140 / 1M $0.030 / 1M
Output Price $1.000 / 1M $0.130 / 1M
Total Input Tokens 130,388 218,731
Output Tokens 40,630 90,507
Reasoning Tokens 786,040 0
Response Time (avg) 112.47s 10.06s
Response Time (max) 950.25s 186.24s
Response Time (total) 2474.28s 221.34s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#133 Qwen3.5-35B-A3B

medium
Cost
$0.009
Time
71.4s
Tokens
8,631 tok

#136 Qwen3.7 Flash

none
Cost
$0.001
Time
34.0s
Tokens
4,814 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Qwen3.5-35B-A3B 5.9 9.3 33.3% 0 206.65s 4,106 23,844 111,462
Qwen3.7 Flash 5.5 10.0 33.3% 0 1.55s 7,911 855 0

Quick Compare

Switch Comparison Pair