Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

Inception: Mercury 2 vs xAI: Grok Build 0.1

Muhtasari

Ulinganisho wa benchmark Mercury 2 vs Grok Build 0.1: average score iko karibu sawa: 7.5 vs 7.6. Mercury 2 ina gharama ya chini ya benchmark: $0.058 vs $0.927. Mercury 2 ni ya haraka zaidi: 2.24s vs 49.90s, na pass rates 54.0% vs 61.9%.

Muundo unaopendekezwa: Mercury 2 - It has the best score here (7.5), while costing about 16.0x less than Grok Build 0.1.

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-06-18

Kipimo Mercury 2 Mercury 2 medium Toleo: 2026-02-24 Grok Build 0.1 Grok Build 0.1 medium Toleo: 2026-05-21
Alama 7.5 7.6
Nafasi #44 #42
Uaminifu 10.0 10.0
Uthabiti 8.8 9.9
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 54.0% 61.9%
Majaribio yasiyo thabiti 3 0
Jumla ya uendeshaji 63 63
Gharama kwa matokeo 0.578 7.124
Jumla ya gharama $0.058 $0.927
Bei ya ingizo $0.250 / 1M $1.000 / 1M
Bei ya toleo $0.750 / 1M $2.000 / 1M
Jumla ya tokeni za ingizo 35,116 44,418
Tokeni za matokeo 4,048 2,782
Tokeni za hoja 61,219 438,018
Muda wa majibu (wastani) 2.24s 49.90s
Muda wa majibu (upeo) 14.63s 252.69s
Muda wa majibu (jumla) 44.72s 1047.92s

Onyesho la kizazi

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#44 Mercury 2

medium
Gharama
$0.002
Muda
2.1s
Tokeni
1,702 tok

#42 xAI: Grok Build 0.1

medium
Gharama
$0.028
Muda
81.3s
Tokeni
14,009 tok

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 6.9 9.9 50.0% 0 1.12s 554 2,546 2,609
Grok Build 0.1 8.3 10.0 75.0% 0 7.43s 2,010 220 12,162
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 8.2 7.7 77.8% 1 2.04s 7,065 296 11,328
Grok Build 0.1 5.7 9.7 33.3% 0 108.46s 8,304 1,138 161,452
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 10.0 10.0 100.0% 0 3.28s 12,909 268 4,887
Grok Build 0.1 10.0 10.0 100.0% 0 32.81s 12,909 231 16,917
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 7.3 5.9 83.3% 1 1.11s 6,234 183 1,656
Grok Build 0.1 10.0 10.0 100.0% 0 10.72s 7,761 180 8,876
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 2.9 7.2 11.1% 1 6.48s 695 41 30,754
Grok Build 0.1 5.3 10.0 33.3% 0 158.00s 1,764 492 175,294
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 4.8 10.0 0.0% 0 821ms 456 137 542
Grok Build 0.1 4.4 9.9 0.0% 0 18.41s 825 76 6,345
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 10.0 10.0 100.0% 0 1.07s 340 14 958
Grok Build 0.1 9.8 10.0 100.0% 0 12.36s 1,362 57 9,599
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 5.4 10.0 33.3% 0 949ms 601 361 2,781
Grok Build 0.1 7.7 10.0 66.7% 0 18.26s 1,689 195 20,841
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 10.0 10.0 100.0% 0 1.89s 6,080 180 1,956
Grok Build 0.1 10.0 10.0 100.0% 0 13.12s 7,263 180 4,969
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Mercury 2 3.0 10.0 0.0% 0 2.58s 182 22 3,748
Grok Build 0.1 3.0 10.0 0.0% 0 53.51s 531 13 21,563

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho