Navigate
Advertise here

DeepSeek V4.1 Flash (high) vs MiMo-V2.6-Pro (medium)

DeepSeek V4.1 Flash (high) leads on average score with 8.7 vs 8.7. MiMo-V2.6-Pro (medium) has the lower benchmark cost at $0.586 vs $0.617. DeepSeek V4.1 Flash (high) is faster at 33.27s vs 148.11s, with pass rates of 75.4% vs 78.3%.

Last updated at: 2026-10-07

Compared models

Rank
#52
Total Output Tokens
464,156
Response Time (avg)
33.27s
Total Cost
$0.617
Rank
#57
Total Output Tokens
523,674
Response Time (avg)
148.11s
Total Cost
$0.586
Recommended model DeepSeek V4.1 Flash (high)

It has the best score here (8.7), while responding about 4.5x faster than MiMo-V2.6-Pro (medium).

Detailed comparison

Metric DeepSeek V4.1 Flash DeepSeek V4.1 Flash high Release: 2026-09-10 MiMo-V2.6-Pro MiMo-V2.6-Pro medium Release: 2026-09-22
Score 8.7 8.7
Rank #52 #57
Reliability 9.7 9.5
Consistency 9.0 7.6
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 75.4% 78.3%
Flaky tests 3 7
Total Runs 69 69
Cost per result 1.843 4.182
Total Cost $0.617 $0.586
Input Price $0.013 / 1M $0.435 / 1M
Output Price $1.320 / 1M $0.870 / 1M
Cache Read Price $0.013 / 1M $0.004 / 1M
Cache Write Price N/A N/A
Total Input Tokens 282,490 298,476
Output Tokens 8,621 7,762
Reasoning Tokens 455,535 515,918
Response Time (avg) 33.27s 148.11s
Response Time (max) 205.53s 904.91s
Response Time (total) 765.12s 3406.53s
Parameters 748B total (16B active) 1.02T total (42B active)
Availability Open source Open source

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#52 DeepSeek V4.1 Flash

high
Cost
$0.035
Time
111.0s
Tokens
29,201 tok

#57 MiMo-V2.6-Pro

medium
Cost
$0.012
Time
185.6s
Tokens
13,493 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
DeepSeek V4.1 Flash 10.0 10.0 100.0% 0 49.19s 7,509 376 108,285
MiMo-V2.6-Pro 8.2 7.2 88.9% 1 150.44s 7,422 785 94,571

Switch Comparison Pair