AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Category

Trivia Ranking

See which AI models perform best on Trivia, which ones stay reliable, and where the biggest gaps appear.

Models Shown

15

Average Trivia Score

3.3

Rank Model Company Trivia Score Score Tests Correct Response Time (avg)
#21 GPT-5.4 medium OpenAI 3.0 8.0 0/1 14.0s
#22 Step 3.7 Flash medium Stepfun 3.0 8.0 0/1 114.0s
#23 GLM 5 Turbo medium Z.ai 3.0 8.0 0/1 40.2s
#24 GPT-5.2 Chat none OpenAI 3.0 7.9 0/1 6.89s
#25 Qwen3.5 Plus 2026-02-15 medium Qwen 3.0 7.9 0/1 103.8s
#26 Qwen3.6 Plus medium Qwen 3.0 7.9 0/1 47.5s
#27 Gemma 4 31B medium Google 3.0 7.8 0/1 90.1s
#28 Gemini 2.5 Flash medium Google 3.0 7.8 0/1 2.76s
#29 Qwen3.5-122B-A10B medium Qwen 3.0 7.8 0/1 52.9s
#30 Qwen3.5-27B medium Qwen 3.0 7.8 0/1 85.1s
#31 DeepSeek V4 Flash high DeepSeek 3.0 7.7 0/1 54.5s
#32 Gemini 3.5 Flash minimal Google 3.0 7.7 0/1 1.76s
#33 Hy3 preview medium Tencent 3.0 7.7 0/1 39.9s
#34 Qwen3.7 Max none Qwen 3.0 7.7 0/1 856ms
#35 Gemini 3 PRO Preview medium Google 3.0 7.6 0/1 0ms

Top Models by Trivia Score

Trivia Score vs Total Cost

Top Models by Response Time (avg)