AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Kategori AI BENCHY

Peringkat Spesifik domain

Lihat model AI mana yang paling baik di Spesifik domain, mana yang tetap andal, dan di mana kesenjangan terbesar muncul. Urutkan berdasarkan: Metrik ↑.

Model yang ditampilkan

15

Rata-rata Skor Spesifik domain

4.8

Model terbaik

GLM 5 Turbo 2.9
Peringkat Model Perusahaan Skor Spesifik domain Skor Tes benar Waktu respons (rata-rata)
#93 Qwen3.6 Plus Preview medium Qwen 3.0 6.3 0/3 22.1s
#98 GLM 5 none Z.ai 3.0 6.1 0/3 2.24s
#106 Grok 4.20 Beta none X AI 3.0 5.8 0/3 611ms
#115 Qwen3.5-27B none Qwen 3.0 5.7 0/3 540ms
#126 gpt-oss-120b none OpenAI 3.0 5.4 0/3 35.0s
#127 Grok 4.20 none X AI 3.0 5.4 0/3 687ms
#130 MiniMax M2.7 medium Minimax 3.0 5.3 0/3 19.0s
#136 Elephant Alpha medium Openrouter 3.0 5.1 0/3 925ms
#137 Elephant Alpha none Openrouter 3.0 5.1 0/3 927ms
#138 Ling-2.6-flash none Inclusionai 3.0 5.0 0/3 4.95s
#143 MiMo-V2.5 none Xiaomi 3.0 4.9 0/3 756ms
#147 GPT-4o-mini none OpenAI 3.0 4.8 0/3 637ms
#154 Qwen3.5-9B none Qwen 3.0 4.6 0/3 464ms
#159 Ling-2.6-1T none Inclusionai 3.0 4.3 0/3 1.04s
#163 Granite 4.1 8B none IBM Granite 3.0 4.0 0/3 357ms

Model teratas menurut Skor Spesifik domain

Skor Spesifik domain vs total biaya

Model teratas menurut Waktu respons (rata-rata)