Datenanalyse und -extraktion: Falsche Antwort
Datenanalyse und -extraktion
Falsche Antwort
Sieh, welche KI-Modelle bei Datenanalyse und -extraktion am ehesten auf Falsche Antwort stoßen, damit du Schwachstellen schneller erkennst. Sortieren nach: Antwortzeit (Durchschnitt) ↑.
Fehlergründe
59/59
Modelle filtern
Keine Modelle entsprechen der aktuellen Suche und den Filtern.
| Rang | Modell | Unternehmen | Falsche Antwort-Anzahl | Kategorie-Score | Gesamtkosten | Korrekte Tests | Antwortzeit (Durchschnitt) |
|---|---|---|---|---|---|---|---|
| #298 | Granite 4.1 8B none | IBM Granite | 2 | 3.0 | $0.007 | 0/2 | 575ms |
| #281 | Mercury 2 none | Inception | 1 | 7.3 | $0.030 | 1/2 | 667ms |
| #311 | LFM2-24B-A2B none | Liquid | 2 | 3.0 | $0.001 | 0/2 | 714ms |
| #290 | Elephant Alpha medium | Openrouter | 1 | 6.5 | $0.000 | 1/2 | 979ms |
| #265 | Mercury 2.5 Preview none | Inception | 1 | 6.3 | $0.018 | 1/2 | 1.04s |
| #288 | Elephant Alpha none | Openrouter | 1 | 6.5 | $0.000 | 1/2 | 1.04s |
| #291 | Granite 4.2 8B none | IBM Granite | 2 | 2.9 | $0.026 | 0/2 | 1.05s |
| #131 | Mercury 2 medium | Inception | 1 | 7.3 | $0.094 | 1/2 | 1.11s |
| #271 | GPT-5.4 Nano none | OpenAI | 1 | 6.5 | $0.041 | 1/2 | 1.11s |
| #260 | Mercury 2.5 Preview low | Inception | 1 | 6.3 | $0.011 | 1/2 | 1.12s |
| #231 | Qwen3.8 27B none | Qwen | 1 | 6.5 | ~$0.006 | 1/2 | 1.14s |
| #254 | Qwen3 Coder Next none | Qwen | 1 | 6.5 | $0.026 | 1/2 | 1.32s |
| #253 | Granite 4.2 8B low | IBM Granite | 1 | 6.5 | $0.012 | 1/2 | 1.39s |
| #309 | Nemotron 3 Nano Omni 30b A3b Reasoning none | NVIDIA | 2 | 3.8 | $0.000 | 0/2 | 1.42s |
| #249 | Granite 4.2 8B medium | IBM Granite | 1 | 6.5 | $0.009 | 1/2 | 1.43s |