Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Google: Gemini 3.5 Flash vs OpenAI: gpt-oss-120b

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-19

Kipimo Gemini 3.5 Flash Gemini 3.5 Flash low Toleo: 2026-05-19 gpt-oss-120b gpt-oss-120b medium Toleo: 2025-08-05 Inapatikana bure
Alama 9.6 5.7
Nafasi #2 #106
Uaminifu 10.0 10.0
Uthabiti 10.0 7.4
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 94.7% 49.1%
Majaribio yasiyo thabiti 0 6
Jumla ya uendeshaji 57 57
Gharama kwa matokeo 1.359 0.152
Jumla ya gharama $0.245 $0.011
Bei ya ingizo $1.500 / 1M $0.000 / 1M
Bei ya toleo $9.000 / 1M $0.000 / 1M
Tokeni za matokeo 2,003 16,594
Tokeni za hoja 20,245 40,637
Muda wa majibu (wastani) 2.84s 16.95s
Muda wa majibu (upeo) 6.44s 50.92s
Muda wa majibu (jumla) 54.00s 203.39s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 2.52s 209 2,536
gpt-oss-120b 6.7 9.9 50.0% 0 10.21s 3,518 2,177
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 5.49s 428 3,146
gpt-oss-120b 4.3 1.1 66.7% 1 26.33s 228 2,549
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 6.44s 351 3,050
gpt-oss-120b 10.0 10.0 100.0% 0 31.18s 694 5,072
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 1.81s 279 1,164
gpt-oss-120b 6.4 5.9 66.7% 1 1.98s 241 1,114
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 7.7 10.0 66.7% 0 3.39s 12 4,538
gpt-oss-120b 2.9 4.4 22.2% 2 50.92s 6,784 20,606
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 2.27s 119 916
gpt-oss-120b 4.3 10.0 0.0% 0 7.90s 107 387
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 9.9 10.0 100.0% 0 1.86s 71 1,652
gpt-oss-120b 9.9 10.0 100.0% 0 7.63s 126 1,799
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 2.35s 288 2,150
gpt-oss-120b 3.2 4.7 22.2% 2 11.80s 1,508 2,092
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 3.27s 234 403
gpt-oss-120b 9.8 10.0 100.0% 0 6.91s 287 1,083
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.5 Flash 10.0 10.0 100.0% 0 1.88s 12 690
gpt-oss-120b 3.0 10.0 0.0% 0 26.51s 3,101 3,758

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho