Navigasi
AI BENCHY
Advertise here

DeepSeek V3.2 vs Nemotron 3.5 Lightning (medium)

Nemotron 3.5 Lightning (medium) unggul dalam skor rata-rata dengan 5.2 vs 5.0. DeepSeek V3.2 memiliki biaya benchmark lebih rendah di $0.054 vs $0.193. DeepSeek V3.2 lebih cepat di 18.25s vs 69.64s, dengan tingkat keberhasilan 37.9% vs 47.0%.

Benchmark dihasilkan dari suite pengujian AI BENCHY pada: 2026-08-11

Peringkat
#218
Total token output
42,097
Waktu respons (rata-rata)
18.25s
Total Biaya
$0.054
Peringkat
#208
Total token output
741,084
Waktu respons (rata-rata)
69.64s
Total Biaya
$0.193
Model yang direkomendasikan DeepSeek V3.2

Its score stays close to the best score here (5.0 vs 5.2), while costing about 3.6x less than Nemotron 3.5 Lightning (medium).

Perbandingan terperinci

Metrik DeepSeek V3.2 DeepSeek V3.2 none Rilis: 2025-12-01 Nemotron 3.5 Lightning Nemotron 3.5 Lightning medium Rilis: 2026-08-11 Tersedia gratis
Skor 5.0 5.2
Peringkat #218 #208
Keandalan 10.0 9.9
Konsistensi 7.7 6.4
Cakupan benchmark 22/22 tes · 66/66 percobaan 22/22 tes · 66/66 percobaan
Tes benar
Tingkat lulus per percobaan 37.9% 47.0%
Tes tidak stabil 6 10
Total Run 66 66
Biaya per hasil 0.870 0.000
Total Biaya $0.054 $0.193
Harga input $0.269 / 1M $0.100 / 1M
Harga output $0.400 / 1M $0.250 / 1M
Total token input 135,780 126,321
Token output 42,097 156,554
Token penalaran 0 584,530
Waktu respons (rata-rata) 18.25s 69.64s
Waktu respons (maks) 115.89s 437.11s
Waktu respons (total) 401.60s 1532.18s

Showcase generasi model

Hamster playing table tennis

Prompt: Create a detailed SVG illustration of a hamster playing table tennis.

#218 DeepSeek V3.2

none
Biaya
$0.002
Waktu
7.0s
Token
1,046 tok

#208 Nemotron 3.5 Lightning

medium
Belum ada hasil showcase yang dihasilkan untuk model ini.
Biaya
$0.000
Waktu
-
Token
0 tok

Model teratas berdasarkan skor

Skor vs Total Biaya

Waktu respons (rata-rata)

Skor vs Waktu respons (rata-rata)

Total token output

Skor vs Total token output

Rincian Kategori

Pemrograman Skor Konsistensi Tingkat lulus per percobaan Tes tidak stabil Tes benar Waktu respons (rata-rata) Token input Token output Token penalaran
DeepSeek V3.2 3.1 6.9 11.1% 1 14.54s 7,279 4,528 0
Nemotron 3.5 Lightning 4.4 5.1 33.3% 2 285.47s 7,623 78,065 297,359

Perbandingan Cepat

Ganti Pasangan Perbandingan