Clasament Respectarea instrucțiunilor
Vezi ce modele AI se descurcă cel mai bine la Respectarea instrucțiunilor, care rămân fiabile și unde apar cele mai mari diferențe.
Modele afișate
15
Media pentru Scor Respectarea instrucțiunilor
8.5
Cel mai bun model
Gemini 3.6 Flash 10.0
311/311
Filtrează modelele
Niciun model nu corespunde căutării și filtrelor curente.
| Rang | Model | Companie | Scor Respectarea instrucțiunilor | Scor | Cost total | Teste corecte | Timp de răspuns (mediu) |
|---|---|---|---|---|---|---|---|
| #223 | Gemma 4 26B A4B none | 6.3 | 5.6 | $0.015 | 1/2 | 690ms | |
| #231 | Qwen3.8 27B none | Qwen | 6.3 | 5.5 | ~$0.006 | 1/2 | 582ms |
| #241 | Inkling Small none | Thinkingmachines | 6.3 | 5.3 | $0.054 | 1/2 | 767ms |
| #248 | Inkling none | Thinkingmachines | 6.3 | 5.2 | $0.147 | 1/2 | 1.72s |
| #254 | Qwen3 Coder Next none | Qwen | 6.3 | 5.1 | $0.026 | 1/2 | 7.78s |
| #263 | GPT-4o-mini none | OpenAI | 6.3 | 5.0 | $0.010 | 1/2 | 1.11s |
| #265 | Mercury 2.5 Preview none | Inception | 6.3 | 4.9 | $0.018 | 1/2 | 1.26s |
| #267 | Nemotron 3 Super none | NVIDIA | 6.3 | 4.8 | $0.008 | 1/2 | 804ms |
| #271 | GPT-5.4 Nano none | OpenAI | 6.3 | 4.8 | $0.041 | 1/2 | 784ms |
| #272 | Trinity Large Thinking high | Arcee AI | 6.3 | 4.8 | $0.592 | 1/2 | 4.12s |
| #283 | Qwen3 Coder Next medium | Qwen | 6.3 | 4.6 | $0.034 | 1/2 | 7.49s |
| #286 | Grok 4.20 Beta none | X AI | 6.3 | 4.4 | $0.087 | 1/2 | 649ms |
| #293 | Grok 4.20 none | X AI | 6.3 | 4.1 | $0.057 | 1/2 | 445ms |
| #311 | LFM2-24B-A2B none | Liquid | 6.3 | 2.2 | $0.001 | 1/2 | 752ms |
| #229 | Qwen3.6 27B none | Qwen | 6.2 | 5.5 | $0.116 | 1/2 | 1.92s |