Navigate
Advertise here

Gemini 3.1 Flash Lite (minimal) vs GPT-5.6 Terra

Gemini 3.1 Flash Lite (minimal) leads on average score with 5.9 vs 5.8. Gemini 3.1 Flash Lite (minimal) has the lower benchmark cost at $0.080 vs $0.377. Gemini 3.1 Flash Lite (minimal) is faster at 3.95s vs 4.11s, with pass rates of 49.3% vs 40.6%.

Last updated at: 2026-10-07

Compared models

Rank
#257
Total Output Tokens
12,875
Response Time (avg)
3.95s
Total Cost
$0.080
Rank
#265
Total Output Tokens
7,381
Response Time (avg)
4.11s
Total Cost
$0.377
Recommended model Gemini 3.1 Flash Lite (minimal)

It has the best score here (5.9), while costing about 4.8x less than GPT-5.6 Terra.

Detailed comparison

Metric Gemini 3.1 Flash Lite Gemini 3.1 Flash Lite minimal Release: 2026-05-08 GPT-5.6 Terra GPT-5.6 Terra none Release: 2026-07-09
Score 5.9 5.8
Rank #257 #265
Reliability 10.0 10.0
Consistency 8.9 9.0
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 49.3% 40.6%
Flaky tests 3 3
Total Runs 69 69
Cost per result 0.791 7.313
Total Cost $0.080 $0.377
Input Price $0.250 / 1M $2.000 / 1M
Output Price $1.500 / 1M $12.000 / 1M
Cache Read Price $0.025 / 1M $0.200 / 1M
Cache Write Price $0.084 / 1M $2.500 / 1M
Total Input Tokens 238,785 213,564
Output Tokens 12,875 7,381
Reasoning Tokens 0 0
Response Time (avg) 3.95s 4.11s
Response Time (max) 50.03s 57.98s
Response Time (total) 90.76s 94.52s
Parameters ~150B total (~10B active) ~1T total (~60B active)
Availability Closed Closed

Cache prices apply to input tokens. Reads reuse cached prompts; writes store them and can cost extra. Output tokens use the output price.

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#257 Gemini 3.1 Flash Lite

minimal
Cost
$0.001
Time
3.7s
Tokens
635 tok

#265 GPT-5.6 Terra

none
Cost
$0.030
Time
14.7s
Tokens
2,082 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Gemini 3.1 Flash Lite 5.5 10.0 33.3% 0 831ms 8,126 666 0
GPT-5.6 Terra 5.5 10.0 33.3% 0 1.00s 7,302 336 0

Switch Comparison Pair