AI BENCHY Categorie
Algemene kennis-ranglijst
Zie welke AI-modellen het best presteren op Algemene kennis, welke betrouwbaar blijven en waar de grootste verschillen zitten. Sorteren op: Metriek โ.
Foutredenen
| Rang | Model | Bedrijf | Algemene kennis-score | Score | Correcte tests | Responstijd (gem.) |
|---|---|---|---|---|---|---|
| #10 | Gemini 3 PRO Preview medium | 0.0 | 8.4 | 0/0 | 0ms | |
| #15 | Qwen3.6 Plus Preview medium | Qwen | 0.0 | 8.2 | 0/0 | 0ms |
| #63 | Laguna M.1 medium | Poolside | 0.0 | 6.9 | 0/0 | 0ms |
| #75 | Laguna Xs.2 medium | Poolside | 0.0 | 6.6 | 0/0 | 0ms |
| #109 | Elephant Alpha medium | Openrouter | 0.0 | 5.5 | 0/0 | 0ms |
| #111 | Nemotron 3 Nano Omni 30b A3b Reasoning medium | NVIDIA | 0.0 | 5.4 | 0/0 | 0ms |
| #115 | Laguna M.1 none | Poolside | 0.0 | 5.4 | 0/0 | 0ms |
| #116 | Elephant Alpha none | Openrouter | 0.0 | 5.3 | 0/0 | 0ms |
| #117 | Laguna Xs.2 none | Poolside | 0.0 | 5.3 | 0/0 | 0ms |
| #134 | Nemotron 3 Nano Omni 30b A3b Reasoning none | NVIDIA | 0.0 | 4.6 | 0/0 | 0ms |
| #138 | Ling-2.6-1T none | Inclusionai | 0.0 | 4.5 | 0/0 | 0ms |
| #4 | GPT-5.5 medium | OpenAI | 2.8 | 8.9 | 0/1 | 37.9s |
| #13 | GPT-5.3-Codex medium | OpenAI | 2.8 | 8.2 | 0/1 | 14.4s |
| #3 | Claude Opus 4.7 medium | Anthropic | 3.0 | 8.9 | 0/1 | 2.25s |
| #5 | Claude Opus 4.7 none | Anthropic | 3.0 | 8.9 | 0/1 | 1.46s |