Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

StepFun: Step 3.5 Flash vs Xiaomi: MiMo-V2-Flash

Last updated at: 2026-04-11

Metric Step 3.5 Flash Step 3.5 Flash none Release: 2026-02-01 MiMo-V2-Flash MiMo-V2-Flash none Release: 2025-12-16
Score 3.0 4.5
Rank #93 #88
Consistency 10.0 7.8
Tests Correct
Attempt pass rate 0.0% 27.8%
Flaky tests 0 5
Total Runs 3 54
Cost per result 0.000 0.753
Total Cost $0.000 $0.023
Input Price $0.100 / 1M $0.090 / 1M
Output Price $0.300 / 1M $0.290 / 1M
Output Tokens 0 68,522
Reasoning Tokens 0 0
Response Time (avg) 0ms 2.79s
Response Time (max) 0ms 19.68s
Response Time (total) 0ms 39.08s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash 3.0 10.0 0.0% 0 0ms 0 0
MiMo-V2-Flash 6.3 3.7 33.3% 1 2.79s 726 0
Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 3.2 8.0 8.3% 1 1.19s 865 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 3.0 10.0 0.0% 0 2.87s 330 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 2.9 5.8 16.7% 1 19.68s 161 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 5.3 7.2 44.4% 1 564ms 24 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 4.6 10.0 0.0% 0 1.67s 104 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 6.5 10.0 50.0% 0 857ms 69 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 3.6 7.2 22.2% 1 1.38s 65,971 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Step 3.5 Flash - - - - - - - -
MiMo-V2-Flash 10.0 10.0 100.0% 0 2.28s 272 0

Quick Compare

Switch Comparison Pair