Kategoria ya AI BENCHY
Orodha ya Uandishi wa msimbo
Ona ni modeli gani za AI zinafanya vizuri zaidi katika Uandishi wa msimbo, zipi zinabaki thabiti, na pengo kubwa liko wapi.
Modeli zilizoonyeshwa
15
Wastani wa Alama ya Uandishi wa msimbo
6.1
Modeli bora
Gemini 3.5 Flash 10.0| Nafasi | Modeli | Kampuni | Alama ya Uandishi wa msimbo | Alama | Majaribio sahihi | Muda wa majibu (wastani) |
|---|---|---|---|---|---|---|
| #50 | Claude Sonnet 4.6 medium | Anthropic | 6.9 | 7.6 | 1/2 | 33.9s |
| #100 | Seed-2.0-Lite none | Bytedance Seed | 6.8 | 5.9 | 1/2 | 2.95s |
| #6 | Gemini 3.5 Flash medium | 6.8 | 9.0 | 1/2 | 9.91s | |
| #67 | GPT-5.4 Nano medium | OpenAI | 6.8 | 7.2 | 1/2 | 21.1s |
| #71 | Seed-2.0-Mini medium | Bytedance Seed | 6.8 | 7.1 | 1/2 | 220.5s |
| #116 | Kimi K2.6 none | Moonshot AI | 6.8 | 5.6 | 1/2 | 122.8s |
| #3 | Gemini 3.5 Flash low | 6.8 | 9.3 | 1/2 | 5.54s | |
| #27 | Qwen3.7 Max none | Qwen | 6.8 | 7.9 | 1/2 | 1.39s |
| #36 | Gemini 3.1 Flash Lite Preview medium | 6.8 | 7.7 | 1/2 | 3.98s | |
| #37 | Gemini 3.1 Flash Lite medium | 6.8 | 7.7 | 1/2 | 3.59s | |
| #41 | Gemini 3 Flash Preview none | 6.8 | 7.7 | 1/2 | 2.19s | |
| #44 | DeepSeek V4 Flash high | DeepSeek | 6.8 | 7.6 | 1/2 | 58.1s |
| #46 | Gemini 3.1 Flash Lite Preview low | 6.8 | 7.6 | 1/2 | 1.56s | |
| #52 | Gemini 3.1 Flash Lite Preview none | 6.8 | 7.5 | 1/2 | 1.06s | |
| #53 | Gemini 3.1 Flash Lite low | 6.8 | 7.4 | 1/2 | 1.71s |