Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

DeepSeek: DeepSeek V3.2 vs Xiaomi: MiMo-V2-Pro

Last updated at: 2026-04-16

Metric DeepSeek V3.2 DeepSeek V3.2 medium Release: 2025-12-01 MiMo-V2-Pro MiMo-V2-Pro medium Release: 2026-03-18
Score 8.0 8.1
Rank #27 #23
Consistency 8.2 8.6
Tests Correct
Attempt pass rate 79.6% 77.8%
Flaky tests 4 3
Total Runs 54 48
Cost per result 0.240 1.320
Total Cost $0.029 $0.159
Input Price $0.260 / 1M $1.000 / 1M
Output Price $0.380 / 1M $3.000 / 1M
Output Tokens 10,620 2,360
Reasoning Tokens 48,511 38,320
Response Time (avg) 46.41s 12.27s
Response Time (max) 180.92s 64.71s
Response Time (total) 835.33s 208.56s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 8.4 9.9 75.0% 0 30.72s 3,773 7,523
MiMo-V2-Pro 10.0 10.0 100.0% 0 3.06s 223 1,107
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 4.7 1.6 66.7% 1 180.92s 626 6,792
MiMo-V2-Pro 10.0 10.0 100.0% 0 52.12s 485 11,361
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 93.11s 571 6,296
MiMo-V2-Pro 4.7 1.6 66.7% 1 64.71s 380 14,186
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 36.09s 207 7,693
MiMo-V2-Pro 7.3 5.8 83.3% 1 17.20s 260 7,484
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 5.3 7.2 44.4% 1 39.32s 3,081 7,856
MiMo-V2-Pro 5.3 10.0 33.3% 0 6.00s 155 1,048
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 5.4 2.5 66.7% 1 31.30s 68 2,366
MiMo-V2-Pro 10.0 10.0 100.0% 0 4.06s 198 424
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 35.78s 1,397 2,845
MiMo-V2-Pro 9.9 10.0 100.0% 0 3.36s 83 667
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 8.2 7.2 88.9% 1 36.87s 390 6,281
MiMo-V2-Pro 7.0 7.2 55.6% 1 4.71s 313 1,179
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 34.81s 507 859
MiMo-V2-Pro 10.0 10.0 100.0% 0 8.19s 263 864

Quick Compare

Switch Comparison Pair