General Intelligence Ranking
See which AI models perform best on General Intelligence, which ones stay reliable, and where the biggest gaps appear. Sort by: Metric ↑.
319/319
Filter models
No models match the current search and filters.
| Rank | Model | Company | General Intelligence Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #106 | Grok Build 0.1 medium | X AI | 4.4 | 7.5 | $1.130 | 0/1 | 18.4s |
| #170 | Qwen3.6 35B A3B medium | Qwen | 4.4 | 6.6 | $0.672 | 0/1 | 8.66s |
| #217 | GPT-5.4 none | OpenAI | 4.4 | 5.8 | $0.397 | 0/1 | 1.78s |
| #247 | DeepSeek V4 Flash 0731 none | DeepSeek | 4.4 | 5.4 | $0.021 | 0/1 | 3.22s |
| #263 | MiMo-V2.5 none | Xiaomi | 4.4 | 5.1 | $0.025 | 0/1 | 6.86s |
| #264 | Qwen3.5-9B none | Qwen | 4.4 | 5.1 | $0.021 | 0/1 | 552ms |
| #307 | Nemotron 3.5 Lightning none | NVIDIA | 4.4 | 4.0 | $0.007 | 0/1 | 586ms |
| #311 | Grok 4.1 Fast none | X AI | 4.4 | 3.8 | $0.008 | 0/1 | 1.08s |
| #65 | GPT-5 Mini medium | OpenAI | 4.5 | 8.1 | $0.241 | 0/1 | 13.5s |
| #98 | GPT-5.4 Nano medium | OpenAI | 4.5 | 7.5 | $0.142 | 0/1 | 4.15s |
| #105 | GPT-5.4 Mini medium | OpenAI | 4.5 | 7.5 | $0.786 | 0/1 | 3.72s |
| #284 | Trinity Large Preview none | Arcee AI | 4.5 | 4.8 | $0.008 | 0/1 | 873ms |
| #146 | Seed-2.0-Code high | Bytedance Seed | 4.6 | 6.9 | $1.273 | 0/1 | 36.4s |
| #242 | Laguna S 2.1 medium | Poolside | 4.6 | 5.4 | $0.053 | 0/1 | 739ms |
| #248 | Laguna S 2.1 high | Poolside | 4.6 | 5.4 | $0.114 | 0/1 | 855ms |