Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

Google: Gemini 3 Flash Preview vs Xiaomi: MiMo-V2.5-Pro

Last updated at: 2026-04-22

Metric Gemini 3 Flash Preview Gemini 3 Flash Preview medium Release: 2025-12-17 MiMo-V2.5-Pro MiMo-V2.5-Pro medium Release: 2026-04-22
Score 10.0 8.1
Rank #1 #23
Consistency 10.0 8.8
Tests Correct
Attempt pass rate 100.0% 75.9%
Flaky tests 0 3
Total Runs 54 54
Cost per result 1.740 1.674
Total Cost $0.314 $0.201
Input Price $0.500 / 1M $1.000 / 1M
Output Price $3.000 / 1M $3.000 / 1M
Output Tokens 2,072 2,735
Reasoning Tokens 97,041 52,571
Response Time (avg) 17.60s 16.17s
Response Time (max) 79.71s 84.22s
Response Time (total) 193.57s 291.09s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 4.13s 305 3,490
MiMo-V2.5-Pro 10.0 10.0 100.0% 0 2.95s 273 1,363
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 79.71s 432 48,771
MiMo-V2.5-Pro 10.0 10.0 100.0% 0 32.58s 543 7,485
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 50.16s 351 12,645
MiMo-V2.5-Pro 10.0 10.0 100.0% 0 53.36s 348 11,870
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 4.72s 279 5,333
MiMo-V2.5-Pro 7.3 5.8 83.3% 1 18.81s 260 8,383
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 21.12s 12 14,908
MiMo-V2.5-Pro 5.3 10.0 33.3% 0 37.87s 275 17,023
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 4.09s 111 1,285
MiMo-V2.5-Pro 5.1 3.3 33.3% 1 4.27s 150 549
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 6.10s 72 4,558
MiMo-V2.5-Pro 9.9 10.0 100.0% 0 2.77s 82 803
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 4.43s 276 4,921
MiMo-V2.5-Pro 6.7 7.9 55.6% 1 5.16s 493 2,187
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Gemini 3 Flash Preview 10.0 10.0 100.0% 0 10.55s 234 1,130
MiMo-V2.5-Pro 10.0 10.0 100.0% 0 16.87s 311 2,908

Quick Compare

Switch Comparison Pair