Urambazaji
AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Compare

ByteDance Seed: Seed-2.0-Lite vs Qwen: Qwen3.5-27B

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-15

Kipimo Seed-2.0-Lite Seed-2.0-Lite medium Toleo: 2026-02-14 Qwen3.5-27B Qwen3.5-27B medium Toleo: 2026-02-24
Nafasi #3 #8
Alama 8.8 8.6
Uthabiti 8.7 9.1
Gharama kwa matokeo 0.870 3.585
Jumla ya gharama $0.105 $0.431
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 87.5% 81.3%
Majaribio yasiyo thabiti 3 2
Jumla ya uendeshaji 48 48
Tokeni za matokeo 2,815 1,658
Tokeni za hoja 44,618 200,786
Muda wa majibu (wastani) 29.39s 52.13s
Muda wa majibu (upeo) 168.71s 163.96s
Muda wa majibu (jumla) 470.29s 834.16s

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Muda wa majibu (wastani)

Alama vs Muda wa majibu (wastani)

Jumla ya tokeni za matokeo

Alama vs Jumla ya tokeni za matokeo

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 23.34s 990 7,037
Qwen3.5-27B 10.0 10.0 100.0% 0 9.69s 102 8,956
Mchanganyiko Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 37.67s 506 4,299
Qwen3.5-27B 10.0 10.0 100.0% 0 163.96s 483 9,991
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 9.07s 246 1,742
Qwen3.5-27B 10.0 10.0 100.0% 0 30.26s 270 16,150
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 5.9 7.2 55.6% 1 88.74s 15 23,897
Qwen3.5-27B 5.3 10.0 33.3% 0 79.53s 43 52,368
Akili ya jumla Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 6.7 3.6 66.7% 1 18.25s 304 1,620
Qwen3.5-27B 6.1 3.1 66.7% 1 101.41s 70 23,147
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 7.26s 71 1,480
Qwen3.5-27B 10.0 10.0 100.0% 0 19.66s 97 11,638
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 9.0 7.9 88.9% 1 11.03s 461 3,532
Qwen3.5-27B 8.2 7.7 77.8% 1 64.61s 245 77,213
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Muda wa majibu (wastani) Tokeni za matokeo Tokeni za hoja
Seed-2.0-Lite 10.0 10.0 100.0% 0 12.38s 222 1,011
Qwen3.5-27B 10.0 10.0 100.0% 0 7.45s 348 1,323

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho