Ranking de Seguimento de instruções
Veja quais modelos de IA vão melhor em Seguimento de instruções, quais permanecem confiáveis e onde aparecem as maiores diferenças.
Modelos exibidos
13
Média de Pontuação de Seguimento de instruções
8.5
Melhor modelo
Gemini 3.6 Flash 10.0
343/343
Filtrar modelos
Nenhum modelo corresponde à pesquisa e aos filtros atuais.
| Posição | Modelo | Empresa | Pontuação de Seguimento de instruções | Pontuação | Custo total | Testes corretos | Tempo de resposta (médio) |
|---|---|---|---|---|---|---|---|
| #291 | Laguna S 2.1 low | Poolside | 4.3 | 5.0 | $0.082 | 0/2 | 423ms |
| #263 | Laguna S 2.1 medium | Poolside | 4.2 | 5.4 | $0.053 | 0/2 | 400ms |
| #271 | Laguna XS 2.1 none | Poolside | 3.8 | 5.3 | $0.008 | 0/2 | 364ms |
| #289 | MiniMax M2.7 medium | Minimax | 3.8 | 5.0 | $0.208 | 0/2 | 12.8s |
| #278 | Granite 4.2 8B medium | IBM Granite | 3.8 | 5.2 | $0.008 | 0/2 | 36.9s |
| #328 | Granite 4.1 8B none | IBM Granite | 3.6 | 4.0 | $0.007 | 0/2 | 344ms |
| #306 | Trinity Large Preview none | Arcee AI | 3.5 | 4.8 | $0.008 | 0/2 | 822ms |
| #327 | Ling 3.0 Tiny none | Inclusionai | 3.1 | 4.0 | $0.000 | 0/2 | 2.44s |
| #282 | Granite 4.2 8B low | IBM Granite | 3.0 | 5.1 | $0.012 | 0/2 | 75.8s |
| #321 | Granite 4.2 8B none | IBM Granite | 3.0 | 4.3 | $0.037 | 0/2 | 79.5s |
| #334 | Grok 4.1 Fast none | X AI | 3.0 | 3.8 | $0.008 | 0/2 | 685ms |
| #284 | Mercury 2.5 low | Inception | 2.9 | 5.1 | $0.011 | 0/2 | 972ms |
| #341 | Jev 1.13 none | Typesafe | 1.5 | 3.2 | $0.003 | 0/1 | 557ms |