Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

DeepSeek: DeepSeek V4 Pro vs OpenAI: gpt-oss-120b

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-24

Kipimo DeepSeek V4 Pro DeepSeek V4 Pro none Toleo: 2026-04-24 gpt-oss-120b gpt-oss-120b medium Toleo: 2025-08-05 Inapatikana bure
Alama 6.7 5.8
Nafasi #59 #76
Uthabiti 9.5 7.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 51.9% 51.9%
Majaribio yasiyo thabiti 1 6
Jumla ya uendeshaji 26 54
Gharama kwa matokeo 0.317 0.144
Jumla ya gharama $0.029 $0.011
Bei ya ingizo $1.740 / 1M $0.000 / 1M
Bei ya toleo $3.480 / 1M $0.000 / 1M
Tokeni za matokeo 1,596 13,493
Tokeni za hoja 0 36,879
Muda wa majibu (wastani) 24.23s 16.08s
Muda wa majibu (upeo) 109.46s 50.92s
Muda wa majibu (jumla) 436.17s 176.88s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 4.8 10.0 25.0% 0 36.12s 221 0
gpt-oss-120b 6.7 9.9 50.0% 0 10.21s 3,518 2,177
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 10.0 10.0 100.0% 0 33.40s 246 0
gpt-oss-120b 4.3 1.1 66.7% 1 26.33s 228 2,549
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 9.5 10.0 100.0% 0 34.55s 826 0
gpt-oss-120b 10.0 10.0 100.0% 0 31.18s 694 5,072
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 10.0 10.0 100.0% 0 54.04s 65 0
gpt-oss-120b 6.4 5.9 66.7% 1 1.98s 241 1,114
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 5.3 10.0 33.3% 0 3.74s 6 0
gpt-oss-120b 2.9 4.4 22.2% 2 50.92s 6,784 20,606
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 4.5 10.0 0.0% 0 6.06s 45 0
gpt-oss-120b 4.3 10.0 0.0% 0 7.90s 107 387
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 6.5 10.0 50.0% 0 3.57s 22 0
gpt-oss-120b 9.9 10.0 100.0% 0 7.63s 126 1,799
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 6.0 7.1 44.4% 1 28.25s 92 0
gpt-oss-120b 3.2 4.7 22.2% 2 11.80s 1,508 2,092
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
DeepSeek V4 Pro 10.0 10.0 100.0% 0 6.47s 73 0
gpt-oss-120b 9.8 10.0 100.0% 0 6.91s 287 1,083

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho