AI BENCHY Category
Trivia Ranking
See which AI models perform best on Trivia, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↓.
Failure Reasons
| Rank | Model | Company | Trivia Score | Score | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|
| #10 | Gemini 3 PRO Preview medium | 0.0 | 8.4 | 0/0 | 0ms | |
| #15 | Qwen3.6 Plus Preview medium | Qwen | 0.0 | 8.2 | 0/0 | 0ms |
| #63 | Laguna M.1 medium | Poolside | 0.0 | 6.9 | 0/0 | 0ms |
| #75 | Laguna Xs.2 medium | Poolside | 0.0 | 6.6 | 0/0 | 0ms |
| #109 | Elephant Alpha medium | Openrouter | 0.0 | 5.5 | 0/0 | 0ms |
| #111 | Nemotron 3 Nano Omni 30b A3b Reasoning medium | NVIDIA | 0.0 | 5.4 | 0/0 | 0ms |
| #115 | Laguna M.1 none | Poolside | 0.0 | 5.4 | 0/0 | 0ms |
| #116 | Elephant Alpha none | Openrouter | 0.0 | 5.3 | 0/0 | 0ms |
| #117 | Laguna Xs.2 none | Poolside | 0.0 | 5.3 | 0/0 | 0ms |
| #134 | Nemotron 3 Nano Omni 30b A3b Reasoning none | NVIDIA | 0.0 | 4.6 | 0/0 | 0ms |
| #138 | Ling-2.6-1T none | Inclusionai | 0.0 | 4.5 | 0/0 | 0ms |