General Intelligence Ranking
See which AI models perform best on General Intelligence, which ones stay reliable, and where the biggest gaps appear.
319/319
Filter models
No models match the current search and filters.
| Rank | Model | Company | General Intelligence Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #306 | Granite 4.1 8B none | IBM Granite | 4.0 | 4.0 | $0.007 | 0/1 | 499ms |
| #318 | Step 3.5 Flash none | Stepfun | 4.0 | 2.3 | $0.020 | 0/1 | 14.4s |
| #319 | LFM2-24B-A2B none | Liquid | 4.0 | 2.2 | $0.001 | 0/1 | 395ms |
| #128 | Qwen3.7 Flash low | Qwen | 4.0 | 7.2 | $0.043 | 0/1 | 14.4s |
| #140 | Grok 4.20 medium | X AI | 3.9 | 7.0 | $0.805 | 0/1 | 24.5s |
| #267 | MiniMax M2.7 medium | Minimax | 3.9 | 5.0 | $0.208 | 0/1 | 38.7s |
| #266 | North Mini Code none | Cohere | 3.9 | 5.1 | $0.000 | 0/1 | 34.8s |
| #211 | Dots 3 Note Preview medium | Dots Studio | 3.8 | 5.9 | $0.000 | 0/1 | 26.2s |
| #279 | GPT-5.4 Nano none | OpenAI | 3.8 | 4.8 | $0.041 | 0/1 | 1.31s |
| #292 | MiniMax M2.5 medium | Minimax | 3.8 | 4.6 | $0.327 | 0/1 | 6.63s |
| #313 | Ling 3.0 Tiny medium | Inclusionai | 3.8 | 3.8 | $0.000 | 0/1 | 31.8s |
| #55 | GPT-5.2 medium | OpenAI | 3.7 | 8.4 | $1.182 | 0/1 | 4.32s |
| #96 | Nemotron 3 Ultra medium | NVIDIA | 3.7 | 7.6 | $0.701 | 0/1 | 2.52s |
| #235 | Nemotron 3.5 Lightning high | NVIDIA | 3.7 | 5.6 | $0.134 | 0/1 | 7.58s |
| #278 | Nemotron 3.5 Lightning low | NVIDIA | 3.7 | 4.8 | $0.140 | 0/1 | 4.61s |