Advertise here
#234

Gemini 3.1 Flash Lite

Google Release: 2026-05-08 Tested on: 2026-05-08 12:04 google/gemini-3.1-flash-lite::high
~150B total (~10B active)MoEClosedEstimated

Summary

Gemini 3.1 Flash Lite scores 5.6 on AI BENCHY and ranks #234. It has 10.0 reliability, a 56.1% pass rate, $2.044 total cost, and 61.96s average response time.

What makes Gemini 3.1 Flash Lite unique: It stands out most in Anti-AI Tricks, where it ranks #3, while Trivia is its weakest area at #13. It uses unusually many reasoning tokens, which can help explain its slower or more expensive runs.

Archived model: this model is no longer updated or tested on new tests.

Identity note

Google: Gemini 3.1 Flash Lite Preview was the preview version of Gemini 3.1 Flash Lite.

Model facts

Researched on 2026-08-12

Estimated
Parameters
~150B total (~10B active)
Architecture
MoE
Availability
Closed
License
-

Best estimate from public evidence; the vendor did not disclose every value. Google does not disclose parameter counts.

Score

5.6

Consistency

6.5

Total Output Tokens

1,357,567

Total Input Tokens

29,134

Input Price

$0.250 / 1M

Output Price

$1.500 / 1M

Tests Correct

Wrong Tests: 8

Attempt pass rate: 56.1%

Flaky tests

4

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

61.96s

Response Time (max): 149.23s

Response Time (total): 1115.31s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#234 Gemini 3.1 Flash Lite

high
No output was saved. The original provider response or failure reason is unavailable.
Cost
$0.000
Time
15.5s
Tokens
0 tok

Price History

Historical pricing data for this model from OpenRouter.

Date Input Price Output Price
2026-06-04 15:40 $0.250 / 1M $1.500 / 1M

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Category Score Consistency Tests Correct
Anti-AI Tricks 8.7 9.5
Coding 3.3 3.3
Combined 5.0 5.0
Data parsing and extraction 10.0 10.0
Domain specific 3.6 7.2
General Intelligence 5.0 2.1
Instructions following 7.3 5.8
Puzzle Solving 5.7 6.8
Tool Calling 10.0 10.0
Trivia 0.0 0.0

Compared models