Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

Mistral: Mistral Small 4 vs OpenAI: GPT-5.4 Nano

Last updated at: 2026-05-01

Metric Mistral Small 4 Mistral Small 4 medium Release: 2026-03-16 GPT-5.4 Nano GPT-5.4 Nano none Release: 2026-03-17
Score 5.7 4.6
Rank #99 #127
Reliability N/A N/A
Consistency 6.8 7.4
Tests Correct
Attempt pass rate 50.0% 33.3%
Flaky tests 7 6
Total Runs 54 54
Cost per result 0.674 0.299
Total Cost $0.034 $0.009
Input Price $0.150 / 1M $0.200 / 1M
Output Price $0.600 / 1M $1.250 / 1M
Output Tokens 15,084 2,762
Reasoning Tokens 39,408 0
Response Time (avg) 5.64s 1.40s
Response Time (max) 30.49s 3.84s
Response Time (total) 101.52s 25.14s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 5.6 3.8 66.7% 3 2.67s 4,055 4,778
GPT-5.4 Nano 3.5 8.0 16.7% 1 1.18s 800 0
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 6.7 3.5 66.7% 1 30.49s 2,796 11,296
GPT-5.4 Nano 7.1 3.7 66.7% 1 1.43s 577 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 3.0 10.0 0.0% 0 25.25s 2,612 10,700
GPT-5.4 Nano 3.0 10.0 0.0% 0 3.84s 280 0
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 7.3 5.9 83.3% 1 1.23s 335 723
GPT-5.4 Nano 6.5 10.0 50.0% 0 1.11s 219 0
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 5.3 7.2 44.4% 1 6.11s 2,621 6,904
GPT-5.4 Nano 2.9 4.4 22.2% 2 926ms 52 0
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 4.8 10.0 0.0% 0 2.05s 821 828
GPT-5.4 Nano 3.8 2.5 33.3% 1 1.31s 180 0
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 7.3 5.8 83.3% 1 1.38s 540 1,031
GPT-5.4 Nano 6.3 10.0 50.0% 0 787ms 84 0
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 3.4 9.7 0.0% 0 2.00s 983 2,338
GPT-5.4 Nano 3.7 7.3 22.2% 1 1.29s 348 0
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Mistral Small 4 10.0 10.0 100.0% 0 3.50s 321 810
GPT-5.4 Nano 10.0 10.0 100.0% 0 3.40s 222 0

Quick Compare

Switch Comparison Pair