Navigate
AI BENCHY
Advertise here

AI BENCHY Compare

DeepSeek: DeepSeek V4 Flash vs Xiaomi: MiMo-V2-Flash

Last updated at: 2026-05-19

Metric DeepSeek V4 Flash DeepSeek V4 Flash none Release: 2026-04-24 Free Available MiMo-V2-Flash MiMo-V2-Flash none Release: 2025-12-16
Score 5.2 4.5
Rank #127 #144
Reliability 10.0 10.0
Consistency 9.2 7.9
Tests Correct
Attempt pass rate 31.6% 26.3%
Flaky tests 2 5
Total Runs 57 57
Cost per result 0.147 0.754
Total Cost $0.008 $0.023
Input Price $0.112 / 1M $0.100 / 1M
Output Price $0.224 / 1M $0.300 / 1M
Output Tokens 4,464 68,534
Reasoning Tokens 0 0
Response Time (avg) 28.01s 2.73s
Response Time (max) 111.96s 19.68s
Response Time (total) 532.17s 40.90s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 20.18s 174 0
MiMo-V2-Flash 3.2 8.0 8.3% 1 1.19s 865 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.3 10.0 0.0% 0 24.04s 471 0
MiMo-V2-Flash 6.3 3.7 33.3% 1 2.79s 726 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.5 2.1 66.7% 1 111.96s 2,664 0
MiMo-V2-Flash 3.0 10.0 0.0% 0 2.87s 330 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 23.79s 195 0
MiMo-V2-Flash 2.9 5.8 16.7% 1 19.68s 161 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 5.3 10.0 33.3% 0 19.73s 18 0
MiMo-V2-Flash 5.3 7.2 44.4% 1 564ms 24 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 4.2 9.9 0.0% 0 23.74s 67 0
MiMo-V2-Flash 4.6 10.0 0.0% 0 1.67s 104 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 6.5 10.0 50.0% 0 17.54s 321 0
MiMo-V2-Flash 6.5 10.0 50.0% 0 857ms 69 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.1 7.3 11.1% 1 22.96s 207 0
MiMo-V2-Flash 3.6 7.2 22.2% 1 1.38s 65,971 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 10.0 10.0 100.0% 0 77.93s 327 0
MiMo-V2-Flash 10.0 10.0 100.0% 0 2.28s 272 0
Trivia Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V4 Flash 3.0 10.0 0.0% 0 3.07s 20 0
MiMo-V2-Flash 3.0 10.0 0.0% 0 1.82s 12 0

Quick Compare

Switch Comparison Pair