Navigate
Advertise here

Ling 3.0 Flash (high) vs GPT-5.6 Terra

The average score is effectively tied at 5.8 vs 5.8. Ling 3.0 Flash (high) has the lower benchmark cost at $0.000 vs $0.516. GPT-5.6 Terra is faster at 4.11s vs 13.08s, with pass rates of 60.9% vs 40.6%.

Last updated at: 2026-10-01

Compared models

Rank
#251
Total Output Tokens
299,456
Response Time (avg)
13.08s
Total Cost
$0.000
Rank
#249
Total Output Tokens
7,381
Response Time (avg)
4.11s
Total Cost
$0.516
Recommended model GPT-5.6 Terra

It has the best score here (5.8), while responding about 3.2x faster than Ling 3.0 Flash (high).

Detailed comparison

Metric Ling 3.0 Flash Ling 3.0 Flash high Release: 2026-07-24 GPT-5.6 Terra GPT-5.6 Terra none Release: 2026-07-09
Score 5.8 5.8
Rank #251 #249
Reliability 9.8 10.0
Consistency 8.6 9.0
Attempts 69/69 69/69
Tests Correct
Attempt pass rate 60.9% 40.6%
Flaky tests 4 3
Total Runs 69 69
Cost per result 0.000 7.313
Total Cost $0.000 $0.516
Input Price $0.021 / 1M $2.000 / 1M
Output Price $0.063 / 1M $12.000 / 1M
Total Input Tokens 132,044 213,564
Output Tokens 92,135 7,381
Reasoning Tokens 207,321 0
Response Time (avg) 13.08s 4.11s
Response Time (max) 75.27s 57.98s
Response Time (total) 287.79s 94.52s
Parameters 124B total (5.1B active) ~1T total (~60B active)
Availability Open source Closed

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#251 Ling 3.0 Flash

high
Cost
$0.000
Time
30.7s
Tokens
10,576 tok

#249 GPT-5.6 Terra

none
Cost
$0.030
Time
14.7s
Tokens
2,082 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
Ling 3.0 Flash 6.7 5.0 66.7% 2 14.85s 8,298 15,602 32,152
GPT-5.6 Terra 5.5 10.0 33.3% 0 1.00s 7,302 336 0

Quick Compare

Switch Comparison Pair