Urambazaji
AI BENCHY
Linganisha Chati
❤️ Made by XCS
Your ad here

AI BENCHY Compare

DeepSeek: DeepSeek V3.2 vs OpenAI: GPT-5 Mini

Linganisha:

Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-03

Kipimo DeepSeek: DeepSeek V3.2 none Toleo: 2025-12-01 OpenAI: GPT-5 Mini medium Toleo: 2025-08-07
Nafasi #39 #33
Wastani wa alama 4.70 5.77
Uthabiti 8.19 8.79
Gharama kwa matokeo 0.132 1.200
Jumla ya gharama $0.007 $0.084
Majaribio sahihi
Kiwango cha kupita kwa kila jaribio 47.6% 57.1%
Majaribio yasiyo thabiti 3 2
Tokeni za matokeo 4,869 4,723
Tokeni za hoja 0 35,392

Modeli bora kwa alama

Alama dhidi ya gharama ya jumla

Mgawanyo wa kategoria

Mbinu za kupinga AI Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 1.00 9.68 0.0% 0 1,411 0
OpenAI: GPT-5 Mini 7.00 9.62 66.7% 0 1,645 5,824
Uchanganuzi na uchimbaji wa data Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 5.38 5.81 66.7% 1 1,710 0
OpenAI: GPT-5 Mini 9.88 10.00 100.0% 0 453 3,200
Mahususi kwa domeni Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 1.00 7.21 22.2% 1 24 0
OpenAI: GPT-5 Mini 1.00 7.21 22.2% 1 293 14,016
Ufuataji wa maagizo Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 8.00 9.99 50.0% 0 66 0
OpenAI: GPT-5 Mini 7.00 6.64 66.7% 1 318 4,992
Puzzle Solving Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 7.67 7.49 88.9% 1 1,136 0
OpenAI: GPT-5 Mini 4.33 9.78 33.3% 0 1,527 5,760
Mwito wa zana Alama Uthabiti Kiwango cha kupita kwa kila jaribio Majaribio yasiyo thabiti Majaribio sahihi Tokeni za matokeo Tokeni za hoja
DeepSeek: DeepSeek V3.2 10.00 10.00 100.0% 0 522 0
OpenAI: GPT-5 Mini 10.00 10.00 100.0% 0 487 1,600

Ulinganisho wa haraka

Badilisha jozi ya ulinganisho