AI BENCHY
Your ad here

AI BENCHY Category

Domain specific Ranking

See which AI models perform best on Domain specific, which ones stay reliable, and where the biggest gaps appear. Sort by: Tests Correct ↓.

Models Shown

8

Average Domain specific Score

4.8

Rank Model Company Domain specific Score Score Tests Correct Response Time (avg)
#85 Elephant none Openrouter 3.0 5.2 0/3 927ms
#86 GPT-5.4 Mini none OpenAI 3.5 5.1 0/3 937ms
#88 Nemotron 3 Super none NVIDIA 3.6 5.1 0/3 6.23s
#89 GPT-4o-mini none OpenAI 3.0 4.9 0/3 637ms
#90 Qwen3.5-9B none Qwen 3.0 4.8 0/3 464ms
#93 GLM 4.7 Flash medium Z.ai 3.5 4.6 0/3 174.6s
#96 GPT-5.4 Nano none OpenAI 2.9 4.5 0/3 926ms
#97 Qwen3.5-9B medium Qwen 3.6 4.4 0/3 137.7s

Top Models by Domain specific Score

Domain specific Score vs Total Cost

Top Models by Response Time (avg)