Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

DeepSeek: DeepSeek V3.2 vs xAI: Grok 4.20

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-10

Kipimo DeepSeek V3.2 DeepSeek V3.2 medium Toleo: 2025-12-01 Grok 4.20 Grok 4.20 medium Toleo: 2026-03-31
Alama 7.2 6.9
Nafasi #61 #68
Uaminifu 10.0 10.0
Uthabiti 7.5 8.3
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 72.8% 63.2%
Majaribio yasiyo thabiti 6 4
Jumla ya uendeshaji 57 57
Gharama kwa matokeo 0.278 7.559
Jumla ya gharama $0.031 $0.756
Bei ya ingizo $0.252 / 1M $1.250 / 1M
Bei ya toleo $0.378 / 1M $2.500 / 1M
Tokeni za matokeo 7,035 1,784
Tokeni za hoja 53,765 128,233
Muda wa majibu (wastani) 46.06s 14.53s
Muda wa majibu (upeo) 180.92s 63.48s
Muda wa majibu (jumla) 875.23s 276.06s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 9.2 10.0 100.0% 0 24.23s 3,247 6,953
Grok 4.20 8.2 7.9 83.3% 1 3.95s 287 8,312
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 4.7 1.6 66.7% 1 180.92s 626 6,792
Grok 4.20 4.3 1.1 66.7% 1 24.33s 250 12,804
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 10.0 10.0 100.0% 0 93.11s 571 6,296
Grok 4.20 10.0 10.0 100.0% 0 17.40s 232 9,556
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 10.0 10.0 100.0% 0 36.09s 207 7,693
Grok 4.20 10.0 10.0 100.0% 0 4.17s 180 5,333
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 2.9 4.4 22.2% 2 24.27s 21 6,838
Grok 4.20 5.3 10.0 33.3% 0 27.03s 375 49,339
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 3.8 2.5 50.0% 1 58.29s 49 2,189
Grok 4.20 3.9 2.6 33.3% 1 24.48s 65 6,440
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 10.0 10.0 100.0% 0 35.78s 1,397 2,845
Grok 4.20 7.3 6.0 83.3% 1 4.42s 40 5,474
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 6.7 5.0 66.7% 2 36.87s 390 6,281
Grok 4.20 7.7 10.0 66.7% 0 6.20s 149 7,913
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 10.0 10.0 100.0% 0 34.81s 507 859
Grok 4.20 3.0 10.0 0.0% 0 13.68s 197 6,620
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V3.2 3.0 10.0 0.0% 0 83.99s 20 7,019
Grok 4.20 3.0 10.0 0.0% 0 63.48s 9 16,442

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho