Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

inclusionAI: Ling-2.6-flash vs Mistral: Mistral Small 4

Summary

Ling-2.6-flash vs Mistral Small 4 benchmark comparison: Mistral Small 4 leads on average score with 5.3 vs 5.0. Ling-2.6-flash has the lower benchmark cost at $0.001 vs $0.068. Ling-2.6-flash is faster at 9.34s vs 9.40s, with pass rates of 31.8% vs 44.4%.

Recommended model: Ling-2.6-flash - Its score stays close to the best score here (5.0 vs 5.3), while costing about 136.1x less than Mistral Small 4.

Last updated at: 2026-06-04

Metric Ling-2.6-flash Ling-2.6-flash none Release: 2026-04-21 Mistral Small 4 Mistral Small 4 medium Release: 2026-03-16
Score 5.0 5.3
Rank #138 #132
Reliability 10.0 10.0
Consistency 9.2 6.9
Tests Correct
Attempt pass rate 31.8% 44.4%
Flaky tests 2 8
Total Runs 63 63
Cost per result 0.005 1.344
Total Cost $0.001 $0.068
Input Price $0.010 / 1M $0.150 / 1M
Output Price $0.030 / 1M $0.600 / 1M
Total Input Tokens 40,718 42,576
Output Tokens 2,878 24,184
Reasoning Tokens 0 84,678
Response Time (avg) 9.34s 9.40s
Response Time (max) 35.34s 59.15s
Response Time (total) 177.48s 197.39s

Generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#138 Ling-2.6-flash

none
No showcase result has been generated for this model yet.
Cost
$0.000
Time
-
Tokens
0 tok

#132 Mistral Small 4

medium
Cost
$0.006
Time
47.9s
Tokens
9,857 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 6.8 8.1 58.3% 1 11.81s 726 573 0
Mistral Small 4 5.6 3.8 66.7% 3 2.67s 708 4,055 4,778
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 5.3 10.0 33.3% 0 11.21s 813 381 0
Mistral Small 4 4.4 5.1 33.3% 2 39.98s 7,636 11,635 54,715
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 3.0 10.0 0.0% 0 35.34s 20,818 1,069 0
Mistral Small 4 3.0 10.0 0.0% 0 25.25s 18,706 2,612 10,700
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 6.5 10.0 50.0% 0 8.48s 8,004 246 0
Mistral Small 4 7.3 5.9 83.3% 1 1.23s 6,171 335 723
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 3.0 10.0 0.0% 0 4.95s 810 24 0
Mistral Small 4 5.3 7.2 44.4% 1 6.11s 742 2,621 6,904
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 4.0 10.0 0.0% 0 1.45s 540 109 0
Mistral Small 4 4.8 10.0 0.0% 0 2.05s 519 821 828
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 9.8 10.0 100.0% 0 5.52s 732 81 0
Mistral Small 4 7.3 5.8 83.3% 1 1.38s 729 540 1,031
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 2.9 7.2 11.1% 1 6.51s 729 151 0
Mistral Small 4 3.4 9.7 0.0% 0 2.17s 735 1,226 2,632
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 3.0 10.0 0.0% 0 18.80s 7,324 229 0
Mistral Small 4 10.0 10.0 100.0% 0 3.50s 6,420 321 810
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling-2.6-flash 3.0 10.0 0.0% 0 1.06s 222 15 0
Mistral Small 4 3.0 10.0 0.0% 0 5.92s 210 18 1,557

Quick Compare

Switch Comparison Pair