Gemini 3.5 FlashvsGPT-5.3-Codex

Gemini 3.5 Flash
GPT-5.3-Codex
API pricingUSD per 1M tokens · lower wins
Input price
$1.50$1.75
Output price
$9.00$14.00
Cached input price
$0.15$0.175
Benchmarks
BullshitBench v2
20%24%
Arena Elo (Code)
14991406
BenchmarksPublished by one model only
Gray Swan IPI · k = 1
14.1%
Gray Swan IPI · k = 10
54.2%
Gray Swan IPI · k = 15
60.5%
SWE-Bench Pro
55.1%
CursorBench v3.2
48.8%
CursorBench v3.1
49.8%
DeepSWE 1.1
37%
MLE-Bench
49.7%
Next.js Evals
83%
Terminal-Bench 2.1
76.2%
MCP Atlas
83.6%
Toolathlon
56.5%
BU Bench
58%
Humanity's Last Exam · no tools
40.2%
ARC-AGI-2
72.1%
OSWorld-Verified
78.4%
Finance Agent v2
57.9%
GDPval-AA
1656
GDPval-AA v2
1349
CharXiv Reasoning
84.2%
MMMU-Pro
83.6%
Blueprint-Bench 2
33.6%
MRCR v2 (8-needle) · 128k average
77.3%
MRCR v2 (8-needle) · 1M pointwise
26.6%
Arena Elo (Text)
1476
Overview
CompanyGoogleOpenAI
Release dateMay 19 2026Feb 5 2026
AccessProprietaryProprietary

Which is better: Gemini 3.5 Flash or GPT-5.3-Codex?

Gemini 3.5 Flash and GPT-5.3-Codex are evenly matched across the 2 benchmarks they both report (BullshitBench v2, Arena Elo (Code)). Gemini 3.5 Flash is cheaper on both input and output: $1.50 vs $1.75 per million input tokens, and $9.00 vs $14.00 per million output tokens. GPT-5.3-Codex shipped 103 days before Gemini 3.5 Flash, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, GPT-5.3-Codex leads at 24% vs Gemini 3.5 Flash at 20%. On Arena Elo (Code), Gemini 3.5 Flash leads at 1499 vs GPT-5.3-Codex at 1406.

Frequently asked questions

Gemini 3.5 Flash was released by Google on May 19 2026.

GPT-5.3-Codex was released by OpenAI on Feb 5 2026.

Gemini 3.5 Flash is cheaper on both input and output: $1.50 vs $1.75 per million input tokens, and $9.00 vs $14.00 per million output tokens. Rates are pay-as-you-go API prices verified on August 18, 2026.

Other comparisons