Urambazaji
AI BENCHY
Linganisha Chati Mbinu
❤️ Made by XCS
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Google: Gemini 3.1 Flash Lite Preview vs xAI: Grok 4.1 Fast

Linganisha:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-06

Kipimo Google: Gemini 3.1 Flash Lite Preview low Toleo: 2026-03-03 xAI: Grok 4.1 Fast medium Toleo: 2025-11-19
Wastani wa alama 7.6 6.4
Nafasi #12 #28
Majaribio sahihi
Uthabiti 10.0 7.8
Gharama kwa matokeo 0.170 0.541
Jumla ya gharama $0.019 $0.049
Kiwango cha kupita kwa kila jaribio 73.3% 71.1%
Majaribio yasiyo thabiti 0 4
common.totalRuns 45 (15 x 3) 45 (15 x 3)
Tokeni za matokeo 1,542 1,056
Tokeni za hoja 6,888 80,419
Muda wa majibu (wastani) 3.49s 27.61s
Muda wa majibu (upeo) 11.91s 121.79s
Muda wa majibu (jumla) 52.29s 220.87s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Wastani wa alama vs Muda wa majibu (wastani)

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 7.0 10.0 66.7% 0 2.18s 456 1,224
xAI: Grok 4.1 Fast 10.0 10.0 100.0% 0 5.65s 102 4,021
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.0 10.0 0.0% 0 11.91s 225 762
xAI: Grok 4.1 Fast 10.0 10.0 100.0% 0 37.64s 261 12,272
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 9.9 10.0 100.0% 0 3.00s 291 696
xAI: Grok 4.1 Fast 9.9 10.0 100.0% 0 6.63s 180 5,409
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 4.0 10.0 33.3% 0 2.36s 18 1,212
xAI: Grok 4.1 Fast 4.0 4.4 66.7% 2 121.79s 11 37,657
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 1.49s 72 753
xAI: Grok 4.1 Fast 5.5 10.0 50.0% 0 5.30s 55 3,489
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 2.76s 243 1,248
xAI: Grok 4.1 Fast 4.0 7.2 44.4% 1 8.08s 187 6,086
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Google: Gemini 3.1 Flash Lite Preview 10.0 10.0 100.0% 0 9.54s 237 993
xAI: Grok 4.1 Fast 10.0 1.6 33.3% 1 27.71s 260 11,485

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho