Urambazaji
AI BENCHY
Your ad here

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Lite vs OpenAI: gpt-oss-120b

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-12

Kipimo Seed-2.0-Lite Seed-2.0-Lite none Toleo: 2026-02-14 gpt-oss-120b gpt-oss-120b medium Toleo: 2025-08-05 Inapatikana bure
Nafasi #45 #43
Wastani wa alama 4.9 5.1
Uthabiti 7.4 7.4
Gharama kwa matokeo 0.214 0.135
Jumla ya gharama $0.015 $0.010
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 56.3% 54.2%
Majaribio yasiyo thabiti 5 5
Jumla ya uendeshaji 48 48
Tokeni za matokeo 2,743 13,210
Tokeni za hoja 0 34,230
Muda wa majibu (wastani) 2.49s 16.65s
Muda wa majibu (upeo) 6.70s 50.92s
Muda wa majibu (jumla) 39.91s 149.88s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Wastani wa alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Wastani wa alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 4.6 22.2% 2 2.93s 703 0
gpt-oss-120b 7.0 9.8 66.7% 0 19.76s 3,463 2,077
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 0.0% 0 6.59s 498 0
gpt-oss-120b 10.0 10.0 100.0% 0 31.18s 694 5,072
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 9.9 10.0 100.0% 0 1.82s 246 0
gpt-oss-120b 5.5 5.9 66.7% 1 1.98s 241 1,114
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 7.2 22.2% 1 1.33s 17 0
gpt-oss-120b 10.0 4.4 22.2% 2 50.92s 6,784 20,606
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 3.45s 294 0
gpt-oss-120b 3.0 10.0 0.0% 0 7.90s 107 387
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 1.06s 73 0
gpt-oss-120b 9.5 10.0 100.0% 0 7.63s 126 1,799
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 4.0 4.4 55.6% 2 2.46s 620 0
gpt-oss-120b 1.7 4.7 22.2% 2 11.80s 1,508 2,092
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 3.94s 292 0
gpt-oss-120b 9.0 10.0 100.0% 0 6.91s 287 1,083

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho