Navigate
AD
Track all your projects in one dashboard. Get 📊stats, 🔥heatmaps and 👀recordings in one self-hosted dashboard.
uxwizz.com

LongCat 2.0 (high) vs Laguna XS 2.1 (medium)

LongCat 2.0 (high) leads on average score with 6.7 vs 6.5. Laguna XS 2.1 (medium) has the lower benchmark cost at $0.068 vs $0.492. Laguna XS 2.1 (medium) is faster at 48.18s vs 153.29s, with pass rates of 51.5% vs 45.5%.

Last updated at: 2026-09-18

Compared models

Rank
#170
Total Output Tokens
386,642
Response Time (avg)
153.29s
Total Cost
$0.492
Rank
#181
Total Output Tokens
522,480
Response Time (avg)
48.18s
Total Cost
$0.068
Recommended model Laguna XS 2.1 (medium)

Its score stays close to the best score here (6.5 vs 6.7), while costing about 7.3x less than LongCat 2.0 (high).

Detailed comparison

Metric LongCat 2.0 LongCat 2.0 high Release: 2026-07-20 Laguna XS 2.1 Laguna XS 2.1 medium Release: 2026-07-02 Free Available
Score 6.7 6.5
Rank #170 #181
Reliability 10.0 10.0
Consistency 8.4 9.2
Attempts 66/66 66/66
Tests Correct
Attempt pass rate 51.5% 45.5%
Flaky tests 4 2
Total Runs 66 66
Cost per result 5.467 0.745
Total Cost $0.492 $0.068
Input Price $0.300 / 1M $0.060 / 1M
Output Price $1.200 / 1M $0.120 / 1M
Total Input Tokens 93,411 118,998
Output Tokens 30,470 30,750
Reasoning Tokens 356,172 491,730
Response Time (avg) 153.29s 48.18s
Response Time (max) 941.66s 422.72s
Response Time (total) 3372.36s 1059.93s
Parameters 1.6T total (48B active) 33B total (3B active)
Availability Open source Weights available

Model generation showcase

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#170 LongCat 2.0

high
Reached the allocated time limit (600 seconds) without receiving showcase output.
Cost
$0.000
Time
600.0s
Tokens
0 tok

#181 Laguna XS 2.1

medium
Cost
$0.001
Time
30.6s
Tokens
4,678 tok

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Input Tokens Output Tokens Reasoning Tokens
LongCat 2.0 7.8 9.3 66.7% 0 505.33s 6,913 244 192,377
Laguna XS 2.1 5.5 10.0 33.3% 0 70.35s 7,995 23,767 83,258

Quick Compare

Switch Comparison Pair