Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite Preview vs OpenAI: GPT-5.3-Codex

Linganisha:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-03

Kipimo Google: Gemini 3.1 Flash Lite Preview high Toleo: 2026-03-03 OpenAI: GPT-5.3-Codex medium Toleo: 2026-02-05
Nafasi #9 #7
Wastani wa alama 7.77 7.93
Uthabiti 9.99 8.84
Gharama kwa matokeo 17.286 4.641
Jumla ya gharama $1.729 $0.465
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 71.4% 78.6%
Majaribio yasiyo thabiti 0 2
Tokeni za matokeo 831 1,201
Tokeni za hoja 1,148,955 30,056

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.00 10.00 100.0% 0 144 193,077
OpenAI: GPT-5.3-Codex 10.00 10.00 100.0% 0 216 1,421
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 9.88 10.00 100.0% 0 279 6,186
OpenAI: GPT-5.3-Codex 10.00 10.00 100.0% 0 234 735
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 4.00 10.00 33.3% 0 18 566,202
OpenAI: GPT-5.3-Codex 4.00 7.21 55.6% 1 64 25,308
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 8.00 9.96 50.0% 0 69 190,053
OpenAI: GPT-5.3-Codex 9.00 10.00 50.0% 0 93 693
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 7.00 10.00 66.7% 0 87 190,953
OpenAI: GPT-5.3-Codex 7.00 7.38 77.8% 1 340 1,407
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.00 10.00 100.0% 0 234 2,484
OpenAI: GPT-5.3-Codex 10.00 10.00 100.0% 0 254 492

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho