Instructions following Ranking
See which AI models perform best on Instructions following, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↑.
316/316
Filter models
No models match the current search and filters.
| Rank | Model | Company | Instructions following Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #139 | Kimi K2.5 medium | Moonshot AI | 10.0 | 7.0 | $0.479 | 2/2 | 92.5s |