AI BENCHY
Advertise here

AI BENCHY Category

Trivia Ranking

See which AI models perform best on Trivia, which ones stay reliable, and where the biggest gaps appear.

Models Shown

15

Average Trivia Score

2.9

Rank Model Company Trivia Score Score Tests Correct Response Time (avg)
#23 Gemini 3.1 Flash Lite Preview medium Google 3.0 8.0 0/1 2.68s
#24 Grok 4.3 medium X AI 3.0 8.0 0/1 44.5s
#25 Gemini 2.5 Flash medium Google 3.0 7.9 0/1 2.76s
#26 GPT-5.4 medium OpenAI 3.0 7.9 0/1 14.0s
#27 Gemini 3.1 Flash Lite medium Google 3.0 7.9 0/1 3.08s
#28 Qwen3.6 Plus medium Qwen 3.0 7.9 0/1 47.5s
#29 Gemini 3 Flash Preview none Google 3.0 7.9 0/1 1.07s
#30 Gemini 3.1 Flash Lite Preview low Google 3.0 7.9 0/1 1.35s
#31 Qwen3.5-122B-A10B medium Qwen 3.0 7.9 0/1 52.9s
#33 Qwen3.5 Plus 2026-04-20 medium Qwen 3.0 7.8 0/1 92.6s
#34 HY3 Preview medium Tencent 3.0 7.8 0/1 39.9s
#35 Claude Sonnet 4.6 medium Anthropic 3.0 7.8 0/1 30.1s
#36 Step 3.5 Flash none Stepfun 3.0 7.8 0/1 114.1s
#37 MiMo-V2-Pro medium Xiaomi 3.0 7.7 0/1 82.7s
#38 Gemma 4 26B A4B medium Google 3.0 7.7 0/1 180.9s

Top Models by Trivia Score

Trivia Score vs Total Cost

Top Models by Response Time (avg)