Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

Google: Gemini 3.1 Pro Preview vs Grok 4.20 Beta

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-26

Kipimo Gemini 3.1 Pro Preview Gemini 3.1 Pro Preview medium Toleo: 2026-02-19 Grok 4.20 Beta Grok 4.20 Beta none Toleo: 2026-03-12
Alama 9.6 5.3
Nafasi #2 #93
Uaminifu Haipo Haipo
Uthabiti 10.0 9.2
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 94.4% 29.6%
Majaribio yasiyo thabiti 0 2
Jumla ya uendeshaji 54 52
Gharama kwa matokeo 3.400 2.255
Jumla ya gharama $0.578 $0.091
Bei ya ingizo $2.000 / 1M $0.000 / 1M
Bei ya toleo $12.000 / 1M $0.000 / 1M
Tokeni za matokeo 1,932 1,591
Tokeni za hoja 40,542 0
Muda wa majibu (wastani) 15.96s 1.19s
Muda wa majibu (upeo) 40.61s 6.48s
Muda wa majibu (jumla) 175.52s 21.37s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 7.90s 112 3,218
Grok 4.20 Beta 4.0 8.4 16.7% 1 597ms 251 0
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 19.88s 405 4,201
Grok 4.20 Beta 5.5 10.0 0.0% 0 1.14s 74 0
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 9.5 10.0 100.0% 0 40.61s 432 9,281
Grok 4.20 Beta 3.0 10.0 0.0% 0 6.48s 282 0
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 7.72s 279 3,904
Grok 4.20 Beta 10.0 10.0 100.0% 0 601ms 197 0
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 7.7 10.0 66.7% 0 32.73s 18 12,424
Grok 4.20 Beta 3.0 10.0 0.0% 0 611ms 160 0
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 11.77s 108 1,179
Grok 4.20 Beta 5.0 10.0 0.0% 0 541ms 87 0
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 9.56s 72 2,236
Grok 4.20 Beta 4.8 10.0 0.0% 0 687ms 60 0
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 7.15s 232 3,117
Grok 4.20 Beta 5.9 7.2 55.6% 1 541ms 291 0
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Gemini 3.1 Pro Preview 10.0 10.0 100.0% 0 23.15s 274 982
Grok 4.20 Beta 10.0 10.0 100.0% 0 4.79s 189 0

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho