Qwen3.5vsGrok 4.6

Qwen3.5
Grok 4.6
Specifications
Parameters
397B
Context window
1M
API pricing
Input price
$0.60$2.00
Output price
$3.60$6.00
Cached input price
$0.50
Cheapest input
$0.08Darkbloom$2.20Amazon Bedrock
Cheapest output
$0.13Darkbloom$6.60Amazon Bedrock
Benchmarks
BullshitBench v2
66%
SWE-Bench Verified
76.4%
SWE-Bench Multilingual
69.3%
DeepSWE 1.1
65.9%
FrontierCode v1.1 (Extended) · extended split
61.3%
APEX-SWE
56.4%
Next.js Evals
71%
Terminal-Bench 4.0
20.3%
Terminal-Bench 3.0
26%
Terminal-Bench 2.0
52.5%
APEX-Agents
57.5%
BrowseComp
69%
Humanity's Last Exam · no tools
28.7%
GPQA Diamond
88.4%
OSWorld-Verified
62.2%
Harvey's Legal Agent Benchmark
15.8%
AA Intelligence Index
61
GDPval-AA v2
1753
AA-Briefcase
1577
CharXiv Reasoning
80.8%
MMMU-Pro
79%
MMMU
85%
threejseval
1528
Overview
CompanyQwenSpaceXAI
Release dateFeb 16 2026Aug 12 2026
AccessOpen WeightProprietary

Other comparisons

Qwen3.5vsClaude Fable 5.1Grok 4.6vsClaude Fable 5.1Qwen3.5vsGPT-6 AstraGrok 4.6vsGPT-6 AstraQwen3.5vsGemini 3.8 FlashGrok 4.6vsGemini 3.8 FlashQwen3.5vsMuse Spark 1.3Grok 4.6vsMuse Spark 1.3Qwen3.5vsDeepSeek-V4.1-FlashGrok 4.6vsDeepSeek-V4.1-FlashQwen3.5vsMistral Medium 3.5Grok 4.6vsMistral Medium 3.5

Frequently asked questions

Qwen3.5 and Grok 4.6 don't publish scores on any of the same benchmarks, so there's no direct head-to-head comparison. Qwen3.5 is cheaper on both input and output: $0.60 vs $2.00 per million input tokens, and $3.60 vs $6.00 per million output tokens. Figures are base-tier rates. Qwen3.5 shipped 177 days before Grok 4.6, so benchmark comparisons should account for the intervening progress.

Qwen3.5 is open weight, while Grok 4.6 is proprietary.

Direct benchmark comparisons are unavailable — Qwen3.5 and Grok 4.6 don't publish scores on any of the same benchmarks.