Advertise here
#34

Pareto

Unbiased Release: 2026-09-18 Tested on: 2026-09-18 02:47 unbiased/pareto::none
-OtherClosed

Summary

Pareto scores 8.9 on AI BENCHY and ranks #34. It has 10.0 reliability, a 86.4% pass rate, $1.172 total cost, and 31.51s average response time.

What makes Pareto unique: It stands out most in Domain specific, where it ranks #1, while Combined is its weakest area at #17.

Model facts

Researched on 2026-09-18

Parameters
-
Architecture
Other
Availability
Closed
License
-

API-only composite model. The vendor describes running multiple models and selecting an answer, rather than a single dense or MoE checkpoint. Constituent models and a meaningful total or active parameter count are not disclosed. OpenRouter exposes no reasoning control for this ID.

Score

8.9

Consistency

9.2

Total Output Tokens

119,230

Total Input Tokens

110,734

Input Price

$2.500 / 1M

Output Price

$7.500 / 1M

Tests Correct

Wrong Tests: 4

Attempt pass rate: 86.4%

Flaky tests

2

Flaky tests had mixed outcomes across runs (at least one pass and one fail).

Response Time (avg)

31.51s

Response Time (max): 143.33s

Response Time (total): 693.29s

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#34 Pareto

none
Cost
$0.034
Time
170.7s
Tokens
4,581 tok

Charts

Choose the first model, then click a second model to open a side-by-side page.

Total Output Tokens

Score vs Total Output Tokens

Quick Compare

Category Breakdown

Category Score Consistency Tests Correct
Anti-AI Tricks 10.0 10.0
Coding 10.0 10.0
Combined 7.2 9.1
Data parsing and extraction 10.0 10.0
Domain specific 7.7 10.0
General Intelligence 10.0 10.0
Instructions following 10.0 10.0
Puzzle Solving 8.2 7.2
Tool Calling 10.0 10.0
Trivia 2.8 1.6

Compared models