Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

Gemini 2.5 Flash (medium) vs Grok 4.7 (low)

Grok 4.7 (low) leads on average score with 7.4 vs 7.3. Gemini 2.5 Flash (medium) has the lower benchmark cost at $0.680 vs $1.933. Gemini 2.5 Flash (medium) is faster at 20.51s vs 63.19s, with pass rates of 68.1% vs 73.9%.

Last updated at: 2026-10-09

Compared models

Rank
#145
Total Output Tokens
241,365
Response Time (avg)
20.51s
Total Cost
$0.680
Rank
#137
Total Output Tokens
255,329
Response Time (avg)
63.19s
Total Cost
$1.933
Recommended model Gemini 2.5 Flash (medium)

Its score stays close to the best score here (7.3 vs 7.4), while costing about 2.8x less than Grok 4.7 (low).

Detailed comparison

Metric Gemini 2.5 Flash Gemini 2.5 Flash medium Release: 2025-06-17 Grok 4.7 Grok 4.7 low Release: 2026-09-21
Score 7.3 7.4
Rank #145 #137
Reliability 10.0 10.0
Consistency 9.6 8.3
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 68.1% 73.9%
Flaky tests 1 5
Total Runs 69 69
Cost per result 4.288 13.804
Total Cost $0.680 $1.933
Input Price $0.300 / 1M $2.000 / 1M
Output Price $2.500 / 1M $6.000 / 1M
Cache Read Price $0.030 / 1M $0.500 / 1M
Cache Write Price $0.084 / 1M N/A
Total Input Tokens 132,507 365,119
Output Tokens 12,739 8,118
Reasoning Tokens 228,626 247,211
Response Time (avg) 20.51s 63.19s
Response Time (max) 140.50s 389.63s
Response Time (total) 471.64s 1453.28s
Parameters ~350B total (~20B active) ~1.7T total (~170B active)
Availability Closed Closed

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#145 Gemini 2.5 Flash

medium
No output was saved. The original provider response or failure reason is unavailable.
Cost
$0.000
Time
274.0s
Tokens
0 tok

#137 SpaceXAI: Grok 4.7

low
Cost
$0.007
Time
15.2s
Tokens
1,537 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 2.5 Flash 7.8 10.0 66.7% 0 41.01s 6,669 543 32,303
Grok 4.7 8.4 7.4 88.9% 1 189.08s 9,579 346 99,053

Switch Comparison Pair