Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Google: Gemini 2.5 Flash vs OpenAI: GPT-5.3-Codex

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-15

Kipimo Gemini 2.5 Flash Gemini 2.5 Flash medium Toleo: 2025-06-17 GPT-5.3-Codex GPT-5.3-Codex medium Toleo: 2026-02-05
Nafasi #15 #5
Alama 8.0 8.7
Uthabiti 9.5 9.1
Gharama kwa matokeo 2.619 4.485
Jumla ya gharama $0.288 $0.539
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 72.9% 83.3%
Majaribio yasiyo thabiti 1 2
Jumla ya uendeshaji 48 48
Tokeni za matokeo 1,370 1,764
Tokeni za hoja 110,522 33,348
Muda wa majibu (wastani) 12.35s 16.59s
Muda wa majibu (upeo) 95.48s 100.93s
Muda wa majibu (jumla) 197.62s 265.39s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 7.8 10.0 66.7% 0 6.98s 249 8,832
GPT-5.3-Codex 10.0 10.0 100.0% 0 4.69s 216 1,421
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 10.0 10.0 100.0% 0 28.44s 303 11,922
GPT-5.3-Codex 10.0 10.0 100.0% 0 19.56s 364 2,731
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 10.0 10.0 100.0% 0 4.06s 279 2,325
GPT-5.3-Codex 10.0 10.0 100.0% 0 3.07s 234 728
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 5.9 7.2 55.6% 1 37.34s 18 80,702
GPT-5.3-Codex 5.9 7.2 55.6% 1 64.31s 64 25,308
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 4.8 10.0 0.0% 0 4.86s 92 1,899
GPT-5.3-Codex 4.6 10.0 0.0% 0 4.87s 187 331
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 9.8 10.0 100.0% 0 2.62s 69 1,203
GPT-5.3-Codex 10.0 10.0 100.0% 0 3.04s 93 693
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 7.7 10.0 66.7% 0 3.94s 126 2,499
GPT-5.3-Codex 9.0 7.9 88.9% 1 5.12s 352 1,644
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 2.5 Flash 10.0 10.0 100.0% 0 6.20s 234 1,140
GPT-5.3-Codex 10.0 10.0 100.0% 0 6.37s 254 492

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho