General Intelligence Ranking
See which AI models perform best on General Intelligence, which ones stay reliable, and where the biggest gaps appear.
319/319
Filter models
No models match the current search and filters.
| Rank | Model | Company | General Intelligence Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #164 | Claude Opus 4.7 none | Anthropic | 10.0 | 6.6 | $0.505 | 1/1 | 3.47s |
| #165 | Gemini 3.5 Flash Lite low | 10.0 | 6.6 | $0.138 | 1/1 | 2.34s | |
| #175 | Hy3 preview medium | Tencent | 10.0 | 6.5 | $0.049 | 1/1 | 16.8s |
| #184 | Mimo V2 PRO medium | Xiaomi | 10.0 | 6.3 | $0.333 | 1/1 | 4.92s |
| #186 | Seed-2.0-Lite none | Bytedance Seed | 10.0 | 6.2 | $0.066 | 1/1 | 3.45s |
| #187 | Gemma 4 31B medium | 10.0 | 6.2 | $0.100 | 1/1 | 9.57s | |
| #193 | Qwen3.7 Flash none | Qwen | 10.0 | 6.1 | $0.019 | 1/1 | 963ms |
| #196 | Inkling low | Thinkingmachines | 10.0 | 6.1 | $0.187 | 1/1 | 3.44s |
| #197 | Trinity Large Thinking medium | Arcee AI | 10.0 | 6.1 | $0.756 | 1/1 | 32.2s |
| #198 | Qwen3.6 Flash none | Qwen | 10.0 | 6.1 | $0.062 | 1/1 | 947ms |
| #199 | Gemma 4 31B none | 10.0 | 6.1 | $0.016 | 1/1 | 2.09s | |
| #202 | Trinity Large Thinking low | Arcee AI | 10.0 | 6.0 | $0.625 | 1/1 | 12.7s |
| #204 | Grok 4.20 Beta medium | X AI | 10.0 | 6.0 | $0.750 | 1/1 | 5.78s |
| #206 | Gemini 3 PRO Preview medium | 10.0 | 6.0 | $0.385 | 1/1 | 9.34s | |
| #208 | Qwen3.5-Flash none | Qwen | 10.0 | 6.0 | $0.073 | 1/1 | 803ms |