Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

DeepSeek: DeepSeek V4 Flash vs Inception: Mercury 2

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-24

Kipimo DeepSeek V4 Flash DeepSeek V4 Flash none Toleo: 2026-04-24 Mercury 2 Mercury 2 medium Toleo: 2026-02-24
Alama 5.3 6.5
Nafasi #89 #62
Uthabiti 9.1 8.6
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 33.3% 53.7%
Majaribio yasiyo thabiti 2 3
Jumla ya uendeshaji 54 54
Gharama kwa matokeo 0.147 0.580
Jumla ya gharama $0.008 $0.047
Bei ya ingizo $0.140 / 1M $0.250 / 1M
Bei ya toleo $0.280 / 1M $0.750 / 1M
Tokeni za matokeo 4,444 3,972
Tokeni za hoja 0 48,333
Muda wa majibu (wastani) 29.39s 2.21s
Muda wa majibu (upeo) 111.96s 14.63s
Muda wa majibu (jumla) 529.10s 37.51s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 3.0 10.0 0.0% 0 20.18s 174 0
Mercury 2 6.9 9.9 50.0% 0 1.12s 2,546 2,609
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 6.3 10.0 0.0% 0 24.04s 471 0
Mercury 2 10.0 10.0 100.0% 0 1.53s 249 2,213
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 4.5 2.1 66.7% 1 111.96s 2,664 0
Mercury 2 10.0 10.0 100.0% 0 3.28s 268 4,887
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 10.0 10.0 100.0% 0 23.79s 195 0
Mercury 2 7.3 5.9 83.3% 1 1.11s 183 1,656
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 5.3 10.0 33.3% 0 19.73s 18 0
Mercury 2 2.9 7.2 11.1% 1 6.48s 41 30,754
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 4.2 9.9 0.0% 0 23.74s 67 0
Mercury 2 4.8 10.0 0.0% 0 821ms 137 542
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 6.5 10.0 50.0% 0 17.54s 321 0
Mercury 2 10.0 10.0 100.0% 0 1.07s 14 958
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 3.1 7.3 11.1% 1 22.96s 207 0
Mercury 2 3.9 7.5 22.2% 1 934ms 354 2,758
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Flash 10.0 10.0 100.0% 0 77.93s 327 0
Mercury 2 10.0 10.0 100.0% 0 1.89s 180 1,956

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho