General Intelligence Ranking
See which AI models perform best on General Intelligence, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↓.
319/319
Filter models
No models match the current search and filters.
| Rank | Model | Company | General Intelligence Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #312 | Qwen3.5-9B medium | Qwen | 2.8 | 3.8 | $0.062 | 0/1 | 226.4s |
| #236 | Hy4 preview low | Tencent | 7.5 | 5.6 | $1.745 | 0/1 | 110.4s |
| #103 | Qwen3.5-27B medium | Qwen | 6.1 | 7.5 | $0.982 | 0/1 | 101.4s |
| #102 | Qwen3.5 Plus 2026-02-15 medium | Qwen | 4.7 | 7.5 | $0.459 | 0/1 | 79.9s |
| #31 | Qwen3.8 Max (0902) high | Qwen | 6.5 | 8.9 | $2.736 | 0/1 | 75.1s |
| #142 | Kimi K2.5 medium | Moonshot AI | 6.5 | 7.0 | $0.479 | 0/1 | 69.7s |
| #95 | Solar Pro 4 xhigh | Upstage | 5.5 | 7.6 | $0.050 | 0/1 | 61.9s |
| #151 | Solar Pro 4 low | Upstage | 5.5 | 6.8 | $0.044 | 0/1 | 61.2s |
| #309 | Ling 3.0 Tiny high | Inclusionai | 5.4 | 3.8 | $0.000 | 0/1 | 60.5s |
| #230 | Owl Alpha medium | Openrouter | 4.3 | 5.6 | $0.000 | 0/1 | 58.6s |
| #136 | DeepSeek V3.2 medium | DeepSeek | 3.4 | 7.0 | $0.078 | 0/1 | 58.3s |
| #276 | Hy4 preview none | Tencent | 5.1 | 4.8 | $0.041 | 0/1 | 58.3s |
| #183 | Ring-2.6-1T medium | Inclusionai | 4.1 | 6.4 | $0.102 | 0/1 | 58.3s |
| #152 | Solar Pro 4 medium | Upstage | 3.0 | 6.8 | $0.047 | 0/1 | 48.0s |
| #227 | Gemini 3.1 Flash Lite high | 5.0 | 5.6 | $2.044 | 0/1 | 45.7s |