AI BENCHY
Your ad here

AI BENCHY Category

Domain specific Ranking

See which AI models perform best on Domain specific, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↑.

Models Shown

15

Average Domain specific Score

4.8

Best Model

GLM 5 3.5
Rank Model Company Domain specific Score Score Tests Correct Response Time (avg)
#17 Gemini 3.1 Flash Lite Preview medium Google 3.0 8.2 0/3 4.21s
#76 Kimi K2.5 none Moonshot AI 5.3 5.5 1/3 4.38s
#23 MiMo-V2-Pro medium Xiaomi 5.3 8.1 1/3 6.00s
#73 Mistral Small 4 medium Mistral 5.3 5.7 1/3 6.11s
#88 Nemotron 3 Super none NVIDIA 3.6 5.1 0/3 6.23s
#54 Mercury 2 medium Inception 2.9 6.5 0/3 6.48s
#12 Gemini 3 PRO Preview medium Google 5.3 8.4 1/3 7.01s
#5 Gemini 3 Flash Preview low Google 5.3 8.8 1/3 8.05s
#50 Hunter Alpha medium OpenRouter 3.0 6.7 0/3 10.5s
#36 GPT-5.3 Chat none OpenAI 3.5 7.7 0/3 13.0s
#51 Nemotron 3 Super medium NVIDIA 2.9 6.7 0/3 16.2s
#8 Qwen3.5 Plus 2026-02-15 medium Qwen 5.3 8.5 1/3 17.5s
#28 GPT-5.2 Chat none OpenAI 5.3 7.9 1/3 17.8s
#80 MiniMax M2.7 medium Minimax 3.0 5.3 0/3 19.0s
#1 Gemini 3 Flash Preview medium Google 10.0 10.0 3/3 21.1s

Top Models by Domain specific Score

Domain specific Score vs Total Cost

Top Models by Response Time (avg)