Instructions following Ranking
See which AI models perform best on Instructions following, which ones stay reliable, and where the biggest gaps appear. Sort by: Total Cost ↑.
316/316
Filter models
No models match the current search and filters.
| Rank | Model | Company | Instructions following Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #278 | Grok 4.20 Multi Agent Beta medium | X AI | 9.8 | 4.8 | $5.599 | 2/2 | 3.52s |