Navigate
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Compared models

Qwen3.8 27B (high) vs Qwen3.8 27B (medium) vs Qwen3.8 27B (low) vs Qwen3.8 27B benchmark comparison: Qwen3.8 27B (high) leads on Score with 8.4. Qwen3.8 27B (low) leads on Reliability with 10.0. Qwen3.8 27B has the lowest Total Cost at ~$0.006. Qwen3.8 27B is fastest at 3.12s.

Last updated at: 2026-08-15

Rank
#42
Total Output Tokens
656,635
Response Time (avg)
167.89s
Total Cost
~$0.287
Rank
#68
Total Output Tokens
136,162
Response Time (avg)
33.05s
Total Cost
~$0.058
Rank
#61
Total Output Tokens
169,329
Response Time (avg)
39.11s
Total Cost
~$0.071
Rank
#214
Total Output Tokens
11,445
Response Time (avg)
3.12s
Total Cost
~$0.006
Recommended model Qwen3.8 27B

It offers the best overall trade-off: a competitive score (5.5), lower cost than the other models in this comparison, and balanced response time.

Detailed comparison

Metric Qwen3.8 27B Qwen3.8 27B high Release: 2026-08-14 Qwen3.8 27B Qwen3.8 27B medium Release: 2026-08-14 Qwen3.8 27B Qwen3.8 27B low Release: 2026-08-14 Qwen3.8 27B Qwen3.8 27B none Release: 2026-08-14
Score 8.4 7.8 7.9 5.5
Rank #42 #68 #61 #214
Reliability 9.6 9.6 10.0 10.0
Consistency 9.2 9.7 10.0 10.0
Attempts 66/66 66/66 66/66 66/66
Tests Correct
Attempt pass rate 77.3% 66.7% 68.2% 40.9%
Flaky tests 2 1 0 0
Total Runs 66 66 66 66
Cost per result ~1.792 ~0.409 ~0.473 ~0.063
Total Cost ~$0.287 ~$0.058 ~$0.071 ~$0.006
Input Price N/A N/A N/A N/A
Output Price N/A N/A N/A N/A
Total Input Tokens 100,179 98,266 99,705 122,463
Output Tokens 904 1,861 1,545 11,445
Reasoning Tokens 655,731 134,301 167,784 0
Response Time (avg) 167.89s 33.05s 39.11s 3.12s
Response Time (max) 640.85s 214.63s 376.04s 44.76s
Response Time (total) 3693.54s 694.04s 860.51s 68.69s
Parameters 27.3B 27.3B 27.3B 27.3B
Availability Weights available Weights available Weights available Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#42 Qwen3.8 27B

high
Cost
~$0.008
Time
270.1s
Tokens
18,963 tok

#68 Qwen3.8 27B

medium
Cost
~$0.002
Time
43.4s
Tokens
2,956 tok

#61 Qwen3.8 27B

low
Cost
~$0.003
Time
92.8s
Tokens
6,902 tok

#214 Qwen3.8 27B

none
Cost
~$0.001
Time
23.9s
Tokens
1,922 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Qwen3.8 27B 8.1 7.0 88.9% 1 248.62s 8,235 346 143,335
Qwen3.8 27B 10.0 10.0 100.0% 0 40.50s 7,893 585 26,808
Qwen3.8 27B 7.7 10.0 66.7% 0 140.92s 8,127 585 86,694
Qwen3.8 27B 5.5 10.0 33.3% 0 1.19s 7,911 408 0

Quick Compare

Switch Comparison Pair

GPT-5.2mediumvsQwen3.8 27BhighSeed 2.1 TurbomediumvsQwen3.8 27BhighQwen3.8 27BhighvsGrok 4.5lowQwen3.8 27BlowvsStep 3.7 FlashmediumGPT-5.2 ChatnonevsQwen3.8 27BlowClaude Opus 4.8lowvsQwen3.8 27BmediumKimi K3maxvsQwen3.8 27BlowNemotron 3.5 LightninghighFree AvailablevsQwen3.8 27BnoneKAT-Coder-Air V2.5mediumvsQwen3.8 27BnoneQwen3.8 27BhighvsGrok 4.6mediumGPT-5.6 TerrahighvsQwen3.8 27BlowSeed-2.0-CodelowvsQwen3.8 27Bmedium