Instructions following Ranking
See which AI models perform best on Instructions following, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↓.
316/316
Filter models
No models match the current search and filters.
| Rank | Model | Company | Instructions following Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #303 | Granite 4.1 8B none | IBM Granite | 3.6 | 4.0 | $0.007 | 0/2 | 344ms |