Navigate
AI BENCHY
Your ad here

AI BENCHY Compare

MoonshotAI: Kimi K2.6 vs NVIDIA: Nemotron 3 Super

Last updated at: 2026-04-20

Metric Kimi K2.6 Kimi K2.6 none Release: 2026-04-20 Nemotron 3 Super Nemotron 3 Super medium Release: 2026-03-11 Free Available
Score 5.8 6.7
Rank #69 #51
Consistency 9.1 8.7
Tests Correct
Attempt pass rate 42.6% 55.6%
Flaky tests 2 3
Total Runs 54 52
Cost per result 0.543 0.000
Total Cost $0.038 $0.000
Input Price $0.950 / 1M $0.090 / 1M
Output Price $4.000 / 1M $0.450 / 1M
Output Tokens 2,973 11,947
Reasoning Tokens 0 29,768
Response Time (avg) 2.05s 19.06s
Response Time (max) 6.65s 87.80s
Response Time (total) 36.93s 305.04s

Top Models by Score

Score vs Total Cost

Response Time (avg)

Score vs Response Time (avg)

Total Output Tokens

Score vs Total Output Tokens

Category Breakdown

Anti-AI Tricks Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 4.6 10.0 25.0% 0 1.39s 471 0
Nemotron 3 Super 10.0 10.0 100.0% 0 10.08s 1,776 3,345
Coding Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 10.0 10.0 100.0% 0 6.65s 1,176 0
Nemotron 3 Super 3.0 10.0 0.0% 0 0ms 0 0
Combined Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 3.0 10.0 0.0% 0 3.38s 290 0
Nemotron 3 Super 10.0 10.0 100.0% 0 87.80s 2,021 9,996
Data parsing and extraction Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 10.0 10.0 100.0% 0 1.32s 201 0
Nemotron 3 Super 10.0 10.0 100.0% 0 18.16s 877 2,607
Domain specific Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 5.3 7.2 44.4% 1 1.48s 42 0
Nemotron 3 Super 2.9 4.4 22.2% 2 16.19s 5,255 6,072
General Intelligence Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 5.4 3.5 33.3% 1 1.55s 138 0
Nemotron 3 Super 3.8 9.9 0.0% 0 27.86s 104 1,149
Instructions following Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 6.5 10.0 50.0% 0 1.64s 72 0
Nemotron 3 Super 7.2 6.5 66.7% 1 7.72s 1,042 2,479
Puzzle Solving Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 3.4 9.7 0.0% 0 1.66s 343 0
Nemotron 3 Super 3.5 9.8 0.0% 0 8.39s 602 2,151
Tool Calling Score Consistency Attempt pass rate Flaky tests Tests Correct Response Time (avg) Output Tokens Reasoning Tokens
Kimi K2.6 10.0 10.0 100.0% 0 4.46s 240 0
Nemotron 3 Super 10.0 10.0 100.0% 0 39.75s 270 1,969

Quick Compare

Switch Comparison Pair