Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite Preview vs xAI: Grok 4.20

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-02

Kipimo Gemini 3.1 Flash Lite Preview Gemini 3.1 Flash Lite Preview low Toleo: 2026-03-03 Grok 4.20 Grok 4.20 medium Toleo: 2026-03-31
Alama 8.0 7.1
Nafasi #20 #40
Uthabiti 10.0 8.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 70.6% 66.7%
Majaribio yasiyo thabiti 0 4
Jumla ya uendeshaji 51 51
Gharama kwa matokeo 0.168 7.358
Jumla ya gharama $0.021 $0.663
Bei ya ingizo $0.250 / 1M $2.000 / 1M
Bei ya toleo $1.500 / 1M $6.000 / 1M
Tokeni za matokeo 1,617 1,494
Tokeni za hoja 7,686 97,078
Muda wa majibu (wastani) 3.28s 9.50s
Muda wa majibu (upeo) 11.91s 29.87s
Muda wa majibu (jumla) 55.80s 161.54s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 8.3 10.0 75.0% 0 2.12s 462 1,638
Grok 4.20 8.2 7.9 83.3% 1 3.36s 280 8,476
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 3.0 10.0 0.0% 0 11.91s 225 762
Grok 4.20 10.0 10.0 100.0% 0 17.40s 232 9,556
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 3.00s 291 696
Grok 4.20 10.0 10.0 100.0% 0 4.17s 180 5,333
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 5.3 10.0 33.3% 0 2.36s 18 1,212
Grok 4.20 5.3 10.0 33.3% 0 27.03s 375 49,339
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 4.0 10.0 0.0% 0 1.54s 69 384
Grok 4.20 5.8 2.8 66.7% 1 7.09s 47 4,252
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.49s 72 753
Grok 4.20 7.3 5.9 83.3% 1 4.42s 40 5,474
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 2.76s 243 1,248
Grok 4.20 6.4 7.7 55.6% 1 3.89s 143 8,028
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 9.54s 237 993
Grok 4.20 3.0 10.0 0.0% 0 13.68s 197 6,620

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho