AI BENCHY
Advertise here

AI BENCHY Category

Domain specific Ranking

See which AI models perform best on Domain specific, which ones stay reliable, and where the biggest gaps appear.

Models Shown

13

Average Domain specific Score

4.8

Rank Model Company Domain specific Score Score Tests Correct Response Time (avg)
#91 GPT-5.5 none OpenAI 2.9 6.4 0/3 1.31s
#103 DeepSeek V4 Pro high DeepSeek 2.9 6.0 0/3 205.7s
#112 GLM 5.1 none Z.ai 2.9 5.7 0/3 1.99s
#133 DeepSeek V3.2 none DeepSeek 2.9 5.2 0/3 4.17s
#149 Nemotron 3 Nano Omni 30b A3b Reasoning medium NVIDIA 2.9 4.6 0/3 56.7s
#23 GLM 5 Turbo medium Z.ai 2.9 8.0 0/3 71.1s
#37 Gemma 4 26B A4B medium Google 2.9 7.6 0/3 23.6s
#72 DeepSeek V3.2 medium DeepSeek 2.9 7.0 0/3 24.3s
#99 gpt-oss-120b medium OpenAI 2.9 6.1 0/3 50.9s
#105 Nemotron 3 Super medium NVIDIA 2.9 5.8 0/3 16.2s
#119 Cobuddy medium Baidu 2.9 5.6 0/3 128.2s
#129 MiniMax M2.5 medium Minimax 2.9 5.3 0/3 237.3s
#148 GPT-5.4 Nano none OpenAI 2.9 4.7 0/3 926ms

Top Models by Domain specific Score

Domain specific Score vs Total Cost

Top Models by Response Time (avg)