Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Anthropic: Claude Opus 4.7 vs OpenAI: GPT-5.4

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-04-16

Kipimo Claude Opus 4.7 Claude Opus 4.7 none Toleo: 2026-04-16 GPT-5.4 GPT-5.4 medium Toleo: 2026-03-05
Alama 9.2 8.2
Nafasi #4 #16
Uthabiti 10.0 8.7
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 88.9% 79.6%
Majaribio yasiyo thabiti 0 3
Jumla ya uendeshaji 54 54
Gharama kwa matokeo 3.155 6.399
Jumla ya gharama $0.505 $0.832
Bei ya ingizo $5.000 / 1M $2.500 / 1M
Bei ya toleo $25.000 / 1M $15.000 / 1M
Tokeni za matokeo 6,326 2,169
Tokeni za hoja 0 48,732
Muda wa majibu (wastani) 3.13s 18.63s
Muda wa majibu (upeo) 18.27s 100.41s
Muda wa majibu (jumla) 56.33s 335.26s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 8.3 10.0 75.0% 0 2.12s 522 0
GPT-5.4 8.3 10.0 75.0% 0 4.11s 240 1,511
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 2.84s 494 0
GPT-5.4 10.0 10.0 100.0% 0 13.03s 389 2,045
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 9.5 10.0 100.0% 0 18.27s 3,504 0
GPT-5.4 10.0 10.0 100.0% 0 20.57s 301 3,543
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 2.15s 324 0
GPT-5.4 10.0 10.0 100.0% 0 5.32s 234 804
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 7.7 10.0 66.7% 0 1.19s 78 0
GPT-5.4 5.3 7.2 44.4% 1 74.27s 61 34,748
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 3.47s 257 0
GPT-5.4 4.7 3.1 33.3% 1 4.92s 145 321
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 1.46s 114 0
GPT-5.4 10.0 10.0 100.0% 0 3.11s 93 897
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 2.58s 661 0
GPT-5.4 8.2 7.2 88.9% 1 9.13s 442 3,832
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Opus 4.7 10.0 10.0 100.0% 0 4.74s 372 0
GPT-5.4 10.0 10.0 100.0% 0 13.28s 264 1,031

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho