Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

DeepSeek: DeepSeek V3.2 vs Qwen: Qwen3.5-122B-A10B

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-03

Kipimo DeepSeek: DeepSeek V3.2 medium Toleo: 2025-12-01 Qwen: Qwen3.5-122B-A10B medium Toleo: 2026-02-24
Nafasi #12 #14
Wastani wa alama 6.98 6.77
Uthabiti 8.75 8.22
Gharama kwa matokeo 0.193 5.137
Jumla ya gharama $0.018 $0.463
Majaribio sahihi 9/14 9/14
Kiwango cha kupita kwa kila jaribio 71.4% 76.2%
Majaribio yasiyo thabiti 2 3
Tokeni za matokeo 6,753 16,751
Tokeni za hoja 30,427 125,394

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 7.00 9.86 66.7% 0 1,171 4,893
Qwen: Qwen3.5-122B-A10B 10.00 10.00 100.0% 0 248 10,486
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 9.88 10.00 100.0% 0 207 7,693
Qwen: Qwen3.5-122B-A10B 9.88 10.00 100.0% 0 270 16,558
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 4.00 7.21 44.4% 1 3,081 7,856
Qwen: Qwen3.5-122B-A10B 1.00 7.21 11.1% 1 15,537 64,889
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 7.00 9.84 50.0% 0 1,397 2,845
Qwen: Qwen3.5-122B-A10B 5.50 5.92 83.3% 1 77 7,372
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 7.00 7.21 88.9% 1 390 6,281
Qwen: Qwen3.5-122B-A10B 7.00 7.21 88.9% 1 297 24,863
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 10.00 10.00 100.0% 0 507 859
Qwen: Qwen3.5-122B-A10B 10.00 10.00 100.0% 0 322 1,226

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho