Navigate
AI BENCHY
Advertise here

Qwen3.7 Flash vs Inkling (low)

The average score is effectively tied at 6.1 vs 6.1. Qwen3.7 Flash has the lower benchmark cost at $0.019 vs $0.187. Inkling (low) is faster at 5.15s vs 10.06s, with pass rates of 36.4% vs 54.6%.

Last updated at: 2026-07-28

Rank
#136
Total Output Tokens
90,507
Response Time (avg)
10.06s
Total Cost
$0.019
Rank
#138
Total Output Tokens
18,922
Response Time (avg)
5.15s
Total Cost
$0.187
Recommended model Qwen3.7 Flash

It has the best score here (6.1), while costing about 10.2x less than Inkling (low).

Detailed comparison

Metric Qwen3.7 Flash Qwen3.7 Flash none Release: 2026-07-28 Inkling Inkling low Release: 2026-07-18
Score 6.1 6.1
Rank #136 #138
Reliability 10.0 9.9
Consistency 9.2 8.6
Tests Correct
Attempt pass rate 36.4% 54.6%
Flaky tests 2 4
Total Runs 66 66
Cost per result 0.262 1.866
Total Cost $0.019 $0.187
Input Price $0.030 / 1M $1.000 / 1M
Output Price $0.130 / 1M $4.050 / 1M
Total Input Tokens 218,731 109,884
Output Tokens 90,507 8,579
Reasoning Tokens 0 10,343
Response Time (avg) 10.06s 5.15s
Response Time (max) 186.24s 41.58s
Response Time (total) 221.34s 113.39s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#136 Qwen3.7 Flash

none
Cost
$0.001
Time
34.0s
Tokens
4,814 tok

#138 Thinking Machines: Inkling

low
Cost
$0.002
Time
6.1s
Tokens
502 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Qwen3.7 Flash 5.5 10.0 33.3% 0 1.55s 7,911 855 0
Inkling 5.1 7.2 22.2% 1 8.71s 7,374 355 4,961

Quick Compare

Switch Comparison Pair