Compare AI models

Specifications
Parameters
30B21B
Context window
—128k
API pricing
Cheapest input
$0.30DeepInfra$0.018Darkbloom
Cheapest output
$1.10Phala$0.09Darkbloom
Benchmarks
SWE-Bench Verified
76%60.7%
Humanity's Last Exam · no tools
22%10.9%
GPQA Diamond
83.5%71.5%
Overview
CompanyMetaOpenAI
Release dateAug 10 2026Aug 5 2025
AccessOpen WeightOpen Weight
Model detailsView modelView model

Frequently asked questions

Muse Glimmer leads gpt-oss-20b on 3 of the 3 benchmarks they both report (SWE-Bench Verified, Humanity's Last Exam, GPQA Diamond). gpt-oss-20b shipped 370 days before Muse Glimmer, so benchmark comparisons should account for the intervening progress.

Muse Glimmer has 30B parameters, while gpt-oss-20b has 21B.

On SWE-Bench Verified, Muse Glimmer leads at 76% vs gpt-oss-20b at 60.7%. On Humanity's Last Exam · no tools, Muse Glimmer leads at 22% vs gpt-oss-20b at 10.9%. On GPQA Diamond, Muse Glimmer leads at 83.5% vs gpt-oss-20b at 71.5%.