Combinat: Răspuns greșit
Combinat
Răspuns greșit
Vezi ce modele AI au cele mai mari șanse să întâmpine Răspuns greșit la Combinat, ca să găsești mai repede punctele slabe. Sortează după: Timp de răspuns (mediu) ↓.
Motive de eșec
90/90
Filtrează modelele
Niciun model nu corespunde căutării și filtrelor curente.
| Rang | Model | Companie | Număr de Răspuns greșit | Scor de categorie | Cost total | Teste corecte | Timp de răspuns (mediu) |
|---|---|---|---|---|---|---|---|
| #296 | Granite 4.2 8B none | IBM Granite | 1 | 3.0 | $0.026 | 0/2 | 424.3s |
| #266 | Laguna S 2.1 low | Poolside | 1 | 3.2 | $0.082 | 0/2 | 412.5s |
| #205 | Qwen3.5-Flash none | Qwen | 1 | 2.9 | $0.073 | 0/2 | 243.6s |
| #156 | KAT-Coder-Pro V2.5 none | Kwaipilot | 1 | 4.1 | $0.487 | 0/2 | 183.1s |
| #306 | Ling 3.0 Tiny high | Inclusionai | 1 | 3.0 | $0.000 | 0/2 | 179.9s |
| #166 | Gemini 3.1 Flash Lite low | 1 | 3.2 | $0.621 | 0/2 | 161.2s | |
| #164 | Gemini 3.1 Flash Lite Preview low | 1 | 3.0 | $0.646 | 0/2 | 160.6s | |
| #310 | Ling 3.0 Tiny medium | Inclusionai | 1 | 3.0 | $0.000 | 0/2 | 139.8s |
| #307 | Ling 3.0 Tiny low | Inclusionai | 1 | 3.0 | $0.000 | 0/2 | 134.8s |
| #223 | Qwen3.5-122B-A10B none | Qwen | 1 | 5.2 | $0.283 | 0/2 | 129.3s |
| #206 | Qwen3.5-35B-A3B none | Qwen | 1 | 3.8 | $0.076 | 0/2 | 128.3s |
| #302 | Ling 3.0 Tiny none | Inclusionai | 1 | 3.0 | $0.000 | 0/2 | 124.2s |
| #197 | Qwen3.5 Plus 2026-04-20 none | Qwen | 1 | 6.4 | $0.122 | 1/2 | 109.7s |
| #190 | Qwen3.7 Flash none | Qwen | 1 | 3.8 | $0.019 | 0/2 | 94.6s |
| #248 | Seed-2.0-Code none | Bytedance Seed | 1 | 3.0 | $0.151 | 0/2 | 78.4s |