- Rank
- #232
- Total Output Tokens
- 760,441
- Response Time (avg)
- 70.11s
- Total Cost
- $0.195
Nemotron 3.5 Lightning (medium) vs GPT-4o-mini
Nemotron 3.5 Lightning (medium) leads on average score with 5.1 vs 5.0. GPT-4o-mini has the lower benchmark cost at $0.010 vs $0.195. GPT-4o-mini is faster at 1.92s vs 70.11s, with pass rates of 48.5% vs 22.7%.
- Rank
- #243
- Total Output Tokens
- 2,913
- Response Time (avg)
- 1.92s
- Total Cost
- $0.010
Recommended model
GPT-4o-mini
Its score stays close to the best score here (5.0 vs 5.1), while costing about 20.0x less than Nemotron 3.5 Lightning (medium).
Detailed comparison
| Metric | Nemotron 3.5 Lightning Nemotron 3.5 Lightning medium Free Available | GPT-4o-mini GPT-4o-mini none |
|---|---|---|
| Score | 5.1 | 5.0 |
| Rank | #232 | #243 |
| Reliability | 9.9 | 10.0 |
| Consistency | 6.0 | 9.9 |
| Attempts | 66/66 | 66/66 |
| Tests Correct | ||
| Attempt pass rate | 48.5% | 22.7% |
| Flaky tests | 11 | 0 |
| Total Runs | 66 | 66 |
| Cost per result | 0.000 | 0.195 |
| Total Cost | $0.195 | $0.010 |
| Input Price | $0.100 / 1M | $0.150 / 1M |
| Output Price | $0.250 / 1M | $0.600 / 1M |
| Total Input Tokens | 126,330 | 53,145 |
| Output Tokens | 165,702 | 2,913 |
| Reasoning Tokens | 594,739 | 0 |
| Response Time (avg) | 70.11s | 1.92s |
| Response Time (max) | 437.11s | 7.58s |
| Response Time (total) | 1542.52s | 30.71s |
| Parameters | 30B total (3B active) | ~8B |
| Availability | Weights available | Closed |
Model generation showcase
Hamster playing table tennis
Prompt: Create a detailed SVG illustration of a hamster playing table tennis.
#232 Nemotron 3.5 Lightning
medium- Cost
- $0.000
- Time
- 130.5s
- Tokens
- 6,750 tok
#243 GPT-4o-mini
none- Cost
- $0.001
- Time
- 6.6s
- Tokens
- 742 tok
Top Models by Score
Score vs Total Cost
Response Time (avg)
Score vs Response Time (avg)
Total Output Tokens
Score vs Total Output Tokens
Category Breakdown
| Anti-AI Tricks | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 6.9 | 7.9 | 66.7% | 1 | 12.34s | 696 | 6,904 | 20,862 | |
| GPT-4o-mini | 4.8 | 10.0 | 25.0% | 0 | 1.34s | 618 | 186 | 0 |
| Coding | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 4.4 | 5.1 | 33.3% | 2 | 285.47s | 7,623 | 78,065 | 297,359 | |
| GPT-4o-mini | 3.2 | 9.6 | 0.0% | 0 | 1.63s | 7,314 | 367 | 0 |
| Combined | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 6.2 | 5.8 | 66.7% | 1 | 136.91s | 98,571 | 15,769 | 112,376 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 6.32s | 29,916 | 1,497 | 0 |
| Data parsing and extraction | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 3.7 | 1.7 | 50.0% | 2 | 23.24s | 7,944 | 4,195 | 30,097 | |
| GPT-4o-mini | 10.0 | 10.0 | 100.0% | 0 | 1.27s | 7,161 | 183 | 0 |
| Domain specific | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 2.9 | 4.4 | 22.2% | 2 | 74.73s | 798 | 54,657 | 91,174 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 743ms | 741 | 17 | 0 |
| General Intelligence | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 5.5 | 10.0 | 0.0% | 0 | 4.36s | 516 | 3,127 | 3,157 | |
| GPT-4o-mini | 4.0 | 10.0 | 0.0% | 0 | 909ms | 480 | 66 | 0 |
| Instructions following | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 8.5 | 6.8 | 83.3% | 1 | 3.75s | 723 | 531 | 4,404 | |
| GPT-4o-mini | 6.3 | 10.0 | 50.0% | 0 | 1.11s | 666 | 69 | 0 |
| Puzzle Solving | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 4.2 | 4.4 | 44.5% | 2 | 10.49s | 726 | 1,170 | 18,195 | |
| GPT-4o-mini | 3.5 | 10.0 | 0.0% | 0 | 1.21s | 651 | 308 | 0 |
| Tool Calling | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 10.0 | 10.0 | 100.0% | 0 | 4.92s | 8,526 | 154 | 2,518 | |
| GPT-4o-mini | 10.0 | 10.0 | 100.0% | 0 | 2.51s | 5,400 | 205 | 0 |
| Trivia | Score | Consistency | Attempt pass rate | Flaky tests | Tests Correct | Response Time (avg) | Input Tokens | Output Tokens | Reasoning Tokens |
|---|---|---|---|---|---|---|---|---|---|
| Nemotron 3.5 Lightning | 3.0 | 10.0 | 0.0% | 0 | 44.02s | 207 | 1,130 | 14,597 | |
| GPT-4o-mini | 3.0 | 10.0 | 0.0% | 0 | 794ms | 198 | 15 | 0 |
Quick Compare
Switch Comparison Pair
Mistral Small 4nonevsNemotron 3.5 LightningmediumFree AvailableNemotron 3.5 LightningmediumFree AvailablevsQwen3 Coder NextnoneGPT-4o-mininonevsLaguna S 2.1lowFree AvailableNemotron 3.5 LightningmediumFree AvailablevsMiMo-V2.5noneNemotron 3.5 LightningmediumFree AvailablevsQwen3.5-9BnoneNemotron 3.5 LightningmediumFree AvailablevsInklingnoneDots 3 Note PreviewnoneFree AvailablevsNemotron 3.5 LightningmediumFree AvailableNorth Mini CodenoneFree AvailablevsNemotron 3.5 LightningmediumFree AvailableMiniMax M2.7mediumvsGPT-4o-mininoneNemotron 3.5 LightningmediumFree AvailablevsQwen3.6 35B A3BnoneNemotron 3.5 LightninglowFree AvailablevsGPT-4o-mininoneLing-2.6-1TnonevsNemotron 3.5 LightningmediumFree Available