Navigate
AI BENCHY
Advertise here

MoonshotAI: Kimi K2.6 vs Qwen: Qwen3.5-Flash

Qwen3.5-Flash leads on average score with 6.1 vs 5.8. Qwen3.5-Flash has the lower benchmark cost at $0.073 vs $0.184. Kimi K2.6 is faster at 19.58s vs 25.28s, with pass rates of 34.9% vs 39.4%.

Recommended modelQwen3.5-FlashIt has the best score here (6.1), while costing about 2.5x less than Kimi K2.6.

Last updated at: 2026-07-20

Metric Kimi K2.6 Kimi K2.6 none Release: 2026-04-20 Qwen3.5-Flash Qwen3.5-Flash none Release: 2026-02-24
Score 5.8 6.1
Rank #138 #125
Reliability 10.0 10.0
Consistency 9.3 9.3
Tests Correct
Attempt pass rate 34.9% 39.4%
Flaky tests 2 2
Total Runs 66 66
Cost per result 3.199 0.933
Total Cost $0.184 $0.073
Input Price $0.684 / 1M $0.065 / 1M
Output Price $3.420 / 1M $0.260 / 1M
Total Input Tokens 116,970 282,347
Output Tokens 30,253 209,201
Reasoning Tokens 0 0
Response Time (avg) 19.58s 25.28s
Response Time (max) 238.89s 480.96s
Response Time (total) 430.85s 556.24s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#138 MoonshotAI: Kimi K2.6

none
Cost
$0.020
Time
127.4s
Tokens
4,429 tok

#125 Qwen3.5-Flash

none
Cost
$0.003
Time
47.4s
Tokens
7,799 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Kimi K2.6 5.5 9.8 33.3% 0 82.57s 5,986 14,754 0
Qwen3.5-Flash 5.5 10.0 33.3% 0 850ms 7,913 519 0

Quick Compare

Switch Comparison Pair