Qwen3vsGrok 4.5

Qwen3
Grok 4.5
Specifications
Parameters
235B
Context window
128k
Benchmarks
Nonsense detection
BullshitBench v2
54%
Prompt injection robustness
Gray Swan IPI · k = 1
13.4%
Prompt injection robustness
Gray Swan IPI · k = 10
54.2%
Prompt injection robustness
Gray Swan IPI · k = 15
60.8%
Agentic coding
SWE-Bench Pro
64.7%
Multilingual coding
SWE-Bench Multilingual
78%
Agentic coding
CursorBench v3.2
66.7%
Agentic coding
DeepSWE 1.1
54%
Agentic coding
DeepSWE 1.0
62%
Next.js coding
Next.js Evals
83%
Agentic computer work
Frontier-Bench v0.1
17.8%
Agentic terminal coding
Terminal-Bench 2.1
83.3%
Agentic legal work
Harvey's Legal Agent Benchmark
12.92%
Medical admin work
MedScribe
86.88%
Community preference (code)
Arena Elo (Code)
1549
Overview
CompanyQwenSpaceXAI
Release dateApr 29 2025Jul 8 2026
AccessOpen WeightProprietary

Which is better: Qwen3 or Grok 4.5?

Qwen3 and Grok 4.5 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Qwen3 shipped 435 days before Grok 4.5, so benchmark comparisons should account for the intervening progress.

Qwen3 is open weight, while Grok 4.5 is proprietary.

Direct benchmark comparisons are unavailable — Qwen3 and Grok 4.5 don't publish scores on any of the same benchmarks.

Frequently asked questions

Qwen3 was released by Qwen on Apr 29 2025.

Grok 4.5 was released by SpaceXAI on Jul 8 2026.

Qwen3 is an open weight model released by Qwen. Grok 4.5 is a proprietary model released by SpaceXAI.

Other comparisons