AI BENCHY
Your ad here

AI BENCHY Category

Domain specific Ranking

See which AI models perform best on Domain specific, which ones stay reliable, and where the biggest gaps appear. Sort by: Response Time (avg) ↓.

Models Shown

15

Average Domain specific Score

4.8

Rank Model Company Domain specific Score Score Tests Correct Response Time (avg)
#86 GPT-5.4 Mini none OpenAI 3.5 5.1 0/3 937ms
#85 Elephant none Openrouter 3.0 5.2 0/3 927ms
#96 GPT-5.4 Nano none OpenAI 2.9 4.5 0/3 926ms
#81 Elephant medium Openrouter 3.0 5.2 0/3 925ms
#59 Qwen3.5-Flash none Qwen 7.7 6.2 2/3 905ms
#78 Trinity Large Preview none Arcee AI 5.3 5.3 1/3 877ms
#74 GLM 4.7 Flash none Z.ai 7.7 5.6 2/3 744ms
#82 Grok 4.20 none X AI 3.0 5.2 0/3 687ms
#92 Qwen3 Coder Next medium Qwen 5.3 4.7 1/3 638ms
#89 GPT-4o-mini none OpenAI 3.0 4.9 0/3 637ms
#79 Grok 4.20 Beta none X AI 3.0 5.3 0/3 611ms
#94 MiMo-V2-Flash none Xiaomi 5.3 4.5 1/3 564ms
#67 Qwen3.5-27B none Qwen 3.0 5.9 0/3 540ms
#91 Mercury 2 none Inception 5.3 4.8 1/3 534ms
#62 Gemini 2.5 Flash none Google 5.9 6.2 1/3 495ms

Top Models by Domain specific Score

Domain specific Score vs Total Cost

Top Models by Response Time (avg)