AI BENCHY
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

AI BENCHY Category

Trivia Ranking

See which AI models perform best on Trivia, which ones stay reliable, and where the biggest gaps appear.

Models Shown

15

Average Trivia Score

2.9

Rank Model Company Trivia Score Score Tests Correct Response Time (avg)
#1 Gemini 3 Flash Preview medium Google 10.0 10.0 1/1 5.50s
#2 Gemini 3.1 Pro Preview medium Google 10.0 9.6 1/1 6.27s
#7 Gemini 3 Flash Preview low Google 10.0 8.8 1/1 2.75s
#3 Claude Opus 4.7 medium Anthropic 3.0 8.9 0/1 2.25s
#5 Claude Opus 4.7 none Anthropic 3.0 8.9 0/1 1.46s
#6 GPT-5.5 low OpenAI 3.0 8.9 0/1 10.1s
#9 Qwen3.6 Max Preview medium Qwen 3.0 8.5 0/1 60.6s
#11 Seed-2.0-Lite medium Bytedance Seed 3.0 8.3 0/1 48.3s
#12 Qwen3.5 Plus 2026-02-15 medium Qwen 3.0 8.2 0/1 103.8s
#14 Gemma 4 31B medium Google 3.0 8.2 0/1 90.1s
#17 Qwen3.5-27B medium Qwen 3.0 8.1 0/1 85.1s
#19 GLM 5 medium Z.ai 3.0 8.1 0/1 67.4s
#20 GLM 5 Turbo medium Z.ai 3.0 8.1 0/1 40.2s
#21 Qwen3.6 35B A3B medium Qwen 3.0 8.0 0/1 32.9s
#22 HY3 Preview high Tencent 3.0 8.0 0/1 47.7s

Top Models by Trivia Score

Trivia Score vs Total Cost

Top Models by Response Time (avg)