Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

Anthropic: Claude Sonnet 4.6 vs OpenAI: GPT-5.3-Codex

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-05-29

Kipimo Claude Sonnet 4.6 Claude Sonnet 4.6 none Toleo: 2026-02-17 GPT-5.3-Codex GPT-5.3-Codex medium Toleo: 2026-02-05
Alama 7.0 8.3
Nafasi #78 #17
Uaminifu 10.0 10.0
Uthabiti 9.7 8.4
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 58.3% 81.7%
Majaribio yasiyo thabiti 1 4
Jumla ya uendeshaji 60 60
Gharama kwa matokeo 2.782 4.887
Jumla ya gharama $0.306 $0.685
Bei ya ingizo $3.000 / 1M $1.750 / 1M
Bei ya toleo $15.000 / 1M $14.000 / 1M
Tokeni za matokeo 9,450 2,336
Tokeni za hoja 0 42,565
Muda wa majibu (wastani) 5.27s 15.95s
Muda wa majibu (upeo) 23.84s 100.93s
Muda wa majibu (jumla) 68.50s 319.08s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 4.8 10.0 25.0% 0 2.94s 1,214 0
GPT-5.3-Codex 8.7 7.9 91.7% 1 4.16s 240 1,722
Uandishi wa msimbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 6.8 10.0 50.0% 0 6.73s 2,112 0
GPT-5.3-Codex 10.0 10.0 100.0% 0 18.45s 514 7,266
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 9.5 10.0 100.0% 0 23.84s 3,766 0
GPT-5.3-Codex 10.0 10.0 100.0% 0 19.56s 364 2,731
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 10.0 10.0 100.0% 0 3.43s 252 0
GPT-5.3-Codex 10.0 10.0 100.0% 0 3.07s 234 728
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 7.7 10.0 66.7% 0 3.54s 413 0
GPT-5.3-Codex 5.9 7.2 55.6% 1 64.31s 64 25,308
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 6.1 3.1 66.7% 1 2.56s 192 0
GPT-5.3-Codex 4.6 10.0 0.0% 0 4.87s 187 331
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 6.5 10.0 50.0% 0 1.96s 90 0
GPT-5.3-Codex 10.0 10.0 100.0% 0 3.04s 93 693
Utatuzi wa mafumbo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 7.7 10.0 66.7% 0 2.53s 533 0
GPT-5.3-Codex 9.0 7.9 88.9% 1 5.05s 356 1,593
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 10.0 10.0 100.0% 0 4.11s 447 0
GPT-5.3-Codex 10.0 10.0 100.0% 0 6.37s 254 492
Maarifa ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Claude Sonnet 4.6 3.0 10.0 0.0% 0 4.67s 431 0
GPT-5.3-Codex 2.8 1.6 33.3% 1 14.43s 30 1,701

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho