Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

DeepSeek: DeepSeek V3.2 vs Xiaomi: MiMo-V2-Flash

Last updated at: 2026-03-15

Metric DeepSeek V3.2 DeepSeek V3.2 medium Release: 2025-12-01 MiMo-V2-Flash MiMo-V2-Flash medium Release: 2025-12-16
Rank #14 #18
Score 8.1 7.9
Consistency 8.5 9.5
Cost per result 0.225 0.316
Total Cost $0.025 $0.035
Tests Correct
Attempt pass rate 79.2% 72.9%
Flaky tests 3 1
Total Runs 48 48
Output Tokens 7,392 11,613
Reasoning Tokens 39,089 106,714
Response Time (avg) 39.48s 25.33s
Response Time (max) 93.11s 96.01s
Response Time (total) 631.71s 253.33s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 7.8 9.9 66.7% 0 33.39s 1,171 4,893
MiMo-V2-Flash 9.9 10.0 100.0% 0 16.79s 1,328 18,739
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 93.11s 571 6,296
MiMo-V2-Flash 9.8 10.0 100.0% 0 75.68s 442 26,859
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 36.09s 207 7,693
MiMo-V2-Flash 6.5 10.0 50.0% 0 0ms 153 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 5.3 7.2 44.4% 1 39.32s 3,081 7,856
MiMo-V2-Flash 5.9 7.2 55.6% 1 96.01s 8,374 42,461
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 5.4 2.5 66.7% 1 31.30s 68 2,366
MiMo-V2-Flash 4.0 10.0 0.0% 0 4.20s 87 488
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 35.78s 1,397 2,845
MiMo-V2-Flash 10.0 10.0 100.0% 0 4.28s 75 3,504
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 8.2 7.2 88.9% 1 36.87s 390 6,281
MiMo-V2-Flash 7.7 10.0 66.7% 0 3.77s 833 1,948
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
DeepSeek V3.2 10.0 10.0 100.0% 0 34.81s 507 859
MiMo-V2-Flash 10.0 10.0 100.0% 0 27.78s 321 12,715

Quick Compare

Switch Comparison Pair