Urambazaji
AI BENCHY
Advertise here

AI BENCHY Compare

Anthropic: Claude Opus 4.8 vs xAI: Grok Build 0.1

Muhtasari

Ulinganisho wa benchmark Claude Opus 4.8 vs Grok Build 0.1: Grok Build 0.1 inaongoza kwa average score: 7.6 vs 7.2. Claude Opus 4.8 ina gharama ya chini ya benchmark: $0.539 vs $0.927. Claude Opus 4.8 ni ya haraka zaidi: 3.47s vs 49.90s, na pass rates 61.9% vs 61.9%.

Muundo unaopendekezwa: Claude Opus 4.8 - Its score stays close to the best score here (7.2 vs 7.6), while costing about 1.7x less than Grok Build 0.1.

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-06-18

Kipimo Claude Opus 4.8 Claude Opus 4.8 none Toleo: 2026-05-28 Grok Build 0.1 Grok Build 0.1 medium Toleo: 2026-05-21
Alama 7.2 7.6
Nafasi #57 #42
Uaminifu 10.0 10.0
Uthabiti 9.2 9.9
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 61.9% 61.9%
Majaribio yasiyo thabiti 2 0
Jumla ya uendeshaji 63 63
Gharama kwa matokeo 4.485 7.124
Jumla ya gharama $0.539 $0.927
Bei ya ingizo $5.000 / 1M $1.000 / 1M
Bei ya toleo $25.000 / 1M $2.000 / 1M
Jumla ya tokeni za ingizo 67,104 44,418
Tokeni za matokeo 8,107 2,782
Tokeni za hoja 0 438,018
Muda wa majibu (wastani) 3.47s 49.90s
Muda wa majibu (upeo) 17.73s 252.69s
Muda wa majibu (jumla) 72.90s 1047.92s

Onyesho la kizazi

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#57 Claude Opus 4.8

none
Gharama
$0.053
Muda
22.0s
Tokeni
2,253 tok

#42 xAI: Grok Build 0.1

medium
Gharama
$0.028
Muda
81.3s
Tokeni
14,009 tok

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 6.5 10.0 50.0% 0 3.40s 834 1,472 0
Grok Build 0.1 8.3 10.0 75.0% 0 7.43s 2,010 220 12,162
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 5.5 10.0 33.3% 0 3.29s 10,590 1,332 0
Grok Build 0.1 5.7 9.7 33.3% 0 108.46s 8,304 1,138 161,452
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 9.5 10.0 100.0% 0 17.73s 29,658 3,259 0
Grok Build 0.1 10.0 10.0 100.0% 0 32.81s 12,909 231 16,917
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 7.3 5.8 83.3% 1 1.77s 10,503 308 0
Grok Build 0.1 10.0 10.0 100.0% 0 10.72s 7,761 180 8,876
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 5.3 7.2 44.4% 1 1.66s 975 61 0
Grok Build 0.1 5.3 10.0 33.3% 0 158.00s 1,764 492 175,294
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 10.0 10.0 100.0% 0 3.48s 708 230 0
Grok Build 0.1 4.4 9.9 0.0% 0 18.41s 825 76 6,345
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 9.9 10.0 100.0% 0 1.37s 909 95 0
Grok Build 0.1 9.8 10.0 100.0% 0 12.36s 1,362 57 9,599
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 7.7 10.0 66.7% 0 2.74s 894 783 0
Grok Build 0.1 7.7 10.0 66.7% 0 18.26s 1,689 195 20,841
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 10.0 10.0 100.0% 0 5.35s 11,775 355 0
Grok Build 0.1 10.0 10.0 100.0% 0 13.12s 7,263 180 4,969
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za ingizo Tokeni za matokeo Tokeni za hoja
Claude Opus 4.8 3.0 10.0 0.0% 0 3.41s 258 212 0
Grok Build 0.1 3.0 10.0 0.0% 0 53.51s 531 13 21,563

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho