General Intelligence Ranking
See which AI models perform best on General Intelligence, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↑.
319/319
Filter models
No models match the current search and filters.
| Rank | Model | Company | General Intelligence Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #252 | Ling-2.6-1T none | Inclusionai | 5.0 | 5.3 | $0.016 | 0/1 | 20.3s |
| #137 | GLM 5.1 medium | Z.ai | 10.0 | 7.0 | $0.727 | 1/1 | 20.9s |
| #210 | Step 3.5 Flash medium | Stepfun | 5.5 | 5.9 | $0.132 | 0/1 | 22.4s |
| #160 | LongCat 2.0 low | Meituan | 3.4 | 6.7 | $0.412 | 0/1 | 22.5s |
| #64 | Qwen3.8 2.4T A95B low | Qwen | 6.1 | 8.1 | $2.672 | 0/1 | 23.0s |
| #288 | Cobuddy medium | Baidu | 4.2 | 4.7 | $0.000 | 0/1 | 23.2s |
| #241 | DeepSeek V4 Flash 0423 none | DeepSeek | 4.2 | 5.4 | $0.040 | 0/1 | 23.7s |
| #140 | Grok 4.20 medium | X AI | 3.9 | 7.0 | $0.805 | 0/1 | 24.5s |
| #134 | Grok 4.3 medium | X AI | 5.4 | 7.0 | $0.741 | 0/1 | 24.7s |
| #219 | North Mini Code medium | Cohere | 5.1 | 5.7 | $0.000 | 0/1 | 25.1s |
| #92 | DeepSeek V4 Flash 0423 high | DeepSeek | 6.1 | 7.6 | $0.042 | 0/1 | 25.2s |
| #127 | Qwen3.5 Plus 2026-04-20 medium | Qwen | 4.9 | 7.2 | $0.323 | 0/1 | 25.3s |
| #78 | Qwen3.7 Plus medium | Qwen | 10.0 | 7.9 | $0.277 | 1/1 | 25.5s |
| #34 | Seed 2.1 Turbo high | Bytedance Seed | 6.5 | 8.9 | $1.797 | 0/1 | 25.7s |
| #211 | Dots 3 Note Preview medium | Dots Studio | 3.8 | 5.9 | $0.000 | 0/1 | 26.2s |