AI BENCHY Compare
Google: Gemini 3 Flash Preview vs OpenAI: GPT-5.2 Chat
Linganisha:
Benchmark zimetengenezwa kutoka seti za majaribio za AI BENCHY tarehe: 2026-03-03
| Kipimo | Google: Gemini 3 Flash Preview low Toleo: 2025-12-17 | OpenAI: GPT-5.2 Chat none Toleo: 2025-12-11 |
|---|---|---|
| Nafasi | #6 | #12 |
| Wastani wa alama | 8.36 | 7.41 |
| Uthabiti | 9.40 | 9.45 |
| Gharama kwa matokeo | 0.602 | 2.261 |
| Jumla ya gharama | $0.067 | $0.227 |
| Majaribio sahihi | ||
| Kiwango cha kupita kwa kila jaribio | 81.0% | 73.8% |
| Majaribio yasiyo thabiti | 1 | 1 |
| Tokeni za matokeo | 1,170 | 14,267 |
| Tokeni za hoja | 18,372 | 0 |
Alama dhidi ya gharama ya jumla
Mgawanyo wa kategoria
| Mbinu za kupinga AI | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 10.00 | 10.00 | 100.0% | 0 | 275 | 2,476 | |
| OpenAI: GPT-5.2 Chat | 10.00 | 10.00 | 100.0% | 0 | 1,651 | 0 |
| Uchanganuzi na uchimbaji wa data | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 10.00 | 10.00 | 100.0% | 0 | 305 | 3,004 | |
| OpenAI: GPT-5.2 Chat | 9.88 | 10.00 | 100.0% | 0 | 980 | 0 |
| Mahususi kwa domeni | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 4.00 | 7.21 | 44.4% | 1 | 12 | 6,410 | |
| OpenAI: GPT-5.2 Chat | 4.00 | 10.00 | 33.3% | 0 | 7,810 | 0 |
| Ufuataji wa maagizo | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 7.50 | 9.99 | 50.0% | 0 | 71 | 2,752 | |
| OpenAI: GPT-5.2 Chat | 5.50 | 6.13 | 66.7% | 1 | 1,528 | 0 |
| Puzzle Solving | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 10.00 | 10.00 | 100.0% | 0 | 273 | 3,315 | |
| OpenAI: GPT-5.2 Chat | 7.00 | 10.00 | 66.7% | 0 | 1,743 | 0 |
| Mwito wa zana | Alama | Uthabiti | Kiwango cha kupita kwa kila jaribio | Majaribio yasiyo thabiti | Majaribio sahihi | Tokeni za matokeo | Tokeni za hoja |
|---|---|---|---|---|---|---|---|
| Google: Gemini 3 Flash Preview | 10.00 | 10.00 | 100.0% | 0 | 234 | 415 | |
| OpenAI: GPT-5.2 Chat | 10.00 | 10.00 | 100.0% | 0 | 555 | 0 |
Ulinganisho wa haraka
Badilisha jozi ya ulinganisho
Claude Sonnet 4.6mediumvsGPT-5.2 ChatnoneGPT-5.2 ChatnonevsGLM 5mediumGemini 3 Flash PreviewlowvsQwen3.5-27BmediumGemini 3 Flash PreviewlowvsQwen3.5 Plus 2026-02-15mediumGemini 2.5 FlashmediumvsGPT-5.2 ChatnoneGemini 3.1 Flash Lite PreviewhighvsGPT-5.2 ChatnoneGPT-5.2 ChatnonevsStep 3.5 FlashmediumInapatikana bureGemini 3 Flash PreviewlowvsGPT-5.3-CodexmediumGemini 3.1 Flash Lite PreviewlowvsGPT-5.2 ChatnoneDeepSeek V3.2mediumvsGPT-5.2 ChatnoneGemini 3.1 Flash Lite PreviewmediumvsGPT-5.2 ChatnoneGPT-5.2 ChatnonevsQwen3.5-122B-A10Bmedium