Agentic Ranking
See which AI models perform best on Agentic, which ones stay reliable, and where the biggest gaps appear. Sort by: Tests Correct ↑.
380/380
Filter models
No models match the current search and filters.
| Rank | Model | Company | Agentic Score | Score | Total Cost | Tests Correct | Response Time (avg) |
|---|---|---|---|---|---|---|---|
| #370 | Laguna Xs.2 none | Poolside | 0.0 | 3.1 | $0.004 | 0/0 | 0ms |
| #371 | Command A+ none | Cohere | 3.0 | 3.0 | $0.000 | 0/1 | 1.05s |
| #372 | Command A+ low | Cohere | 2.8 | 3.0 | $0.073 | 0/1 | 128.5s |
| #373 | Nemotron 3 Nano Omni 30b A3b Reasoning medium | NVIDIA | 0.0 | 2.7 | $0.000 | 0/0 | 0ms |
| #374 | Nemotron 3 Nano Omni 30b A3b Reasoning none | NVIDIA | 0.0 | 2.6 | $0.000 | 0/0 | 0ms |
| #375 | Jev 1.13 none | Typesafe | 0.0 | 2.6 | $0.003 | 0/0 | 0ms |
| #376 | Kev 4B none | Jaredpalmer | 0.0 | 2.4 | $0.002 | 0/0 | 0ms |
| #377 | Step 3.5 Flash none | Stepfun | 0.0 | 1.9 | $0.020 | 0/0 | 0ms |
| #379 | LFM2-24B-A2B none | Liquid | 0.0 | 1.8 | $0.001 | 0/0 | 0ms |
| #380 | Claude Opus 5.5 none | Anthropic | 0.0 | 1.5 | $0.015 | 0/0 | 0ms |
| #1 | Gemini 3.6 Flash high | 10.0 | 9.9 | $1.512 | 1/1 | 65.6s | |
| #2 | Gemini 3.8 Flash high | 10.0 | 9.9 | $1.927 | 1/1 | 115.5s | |
| #3 | GPT-6 Sol high | OpenAI | 10.0 | 9.9 | $0.921 | 1/1 | 53.3s |
| #4 | GPT-6.1 Sol max | OpenAI | 10.0 | 9.9 | $1.006 | 1/1 | 80.2s |
| #5 | GPT-6 Astra max | OpenAI | 10.0 | 9.9 | $5.044 | 1/1 | 69.6s |