领域专项 排名
看看哪些 AI 模型在 领域专项 上表现最好,哪些更稳定,以及差距主要出现在哪里。 排序方式: 测试正确 ↓.
316/316
筛选模型
没有模型匹配当前搜索和筛选条件。
| 排名 | 模型 | 公司 | 领域专项 得分 | 分数 | 总成本 | 测试正确 | 响应时间(平均) |
|---|---|---|---|---|---|---|---|
| #250 | Gemini 3.1 Flash Lite Preview high | 5.3 | 5.3 | $2.310 | 1/3 | 127.6s | |
| #253 | Inkling none | Thinkingmachines | 5.3 | 5.2 | $0.147 | 1/3 | 1.18s |
| #254 | Granite 4.2 8B medium | IBM Granite | 5.3 | 5.2 | $0.009 | 1/3 | 14.0s |
| #256 | Mistral Small 4 none | Mistral | 5.3 | 5.1 | $0.022 | 1/3 | 377ms |
| #257 | Mistral Small 4 medium | Mistral | 5.9 | 5.1 | $0.097 | 1/3 | 6.33s |
| #258 | Granite 4.2 8B low | IBM Granite | 5.3 | 5.1 | $0.012 | 1/3 | 26.2s |
| #259 | Qwen3 Coder Next none | Qwen | 5.3 | 5.1 | $0.026 | 1/3 | 871ms |
| #262 | GLM 5 Turbo none | Z.ai | 5.3 | 5.1 | $0.047 | 1/3 | 1.97s |
| #273 | Hy4 preview none | Tencent | 6.8 | 4.8 | $0.041 | 1/3 | 25.6s |
| #274 | Ring-2.6-1T none | Inclusionai | 5.3 | 4.8 | $0.026 | 1/3 | 56.8s |
| #280 | GLM 4.7 Flash none | Z.ai | 5.3 | 4.8 | $0.016 | 1/3 | 2.22s |
| #281 | Trinity Large Preview none | Arcee AI | 5.3 | 4.8 | $0.008 | 1/3 | 877ms |
| #283 | Grok 4.1 Fast medium | X AI | 5.8 | 4.7 | $0.069 | 1/3 | 121.8s |
| #284 | Laguna M.1 medium | Poolside | 5.3 | 4.7 | $0.033 | 1/3 | 24.1s |
| #286 | Mercury 2 none | Inception | 5.9 | 4.7 | $0.030 | 1/3 | 500ms |