Claude 3.7 Sonnet
Released
Claude 3.7 Sonnet is an AI model released by Anthropic on Monday, Feb 24 2025, 125 days after Claude 3.5 Sonnet (upgraded). Benchmark results (shown below) cover BullshitBench v2, SWE-Bench Verified, and GPQA Diamond.
Benchmarks
Coding
Reasoning & science
Robustness
About
Claude 3.7 Sonnet, released February 24, 2025, was Anthropic's first hybrid reasoning model: a single model that could answer instantly or engage an extended-thinking mode where it reasons step by step before responding, with the thinking budget under API control. It launched alongside the first research preview of Claude Code, Anthropic's agentic command-line coding tool.
The reasoning upgrade showed up most clearly in software engineering, where Claude 3.7 Sonnet scored 62.3% on SWE-Bench Verified — up from 49.0% for its predecessor and the best published score of any model at the time. GPQA Diamond rose to 68.0%. It held the top of the Claude lineup for three months until Claude Sonnet 4 and Opus 4 arrived in May 2025.
Compare Claude 3.7 Sonnet with
Suggested comparisons
Frequently asked questions
Claude 3.7 Sonnet was released by Anthropic on Monday, Feb 24 2025.
Claude 3.7 Sonnet was built by Anthropic. AI safety company building the Claude family of models. Founded in 2021 by former OpenAI researchers.
Claude 3.7 Sonnet reports 3 tracked benchmark scores — BullshitBench v2: 49%; SWE-Bench Verified: 62.3%; GPQA Diamond: 68%. Scores are the figures published at release by Anthropic.
No. Claude 3.7 Sonnet is a proprietary model. The weights are not published — it is available only through the provider's own API, apps, or partner platforms.
Anthropic's previous tracked release was Claude 3.5 Sonnet (upgraded) on Oct 22 2024, 125 days earlier. It was followed by Claude Sonnet 4 on May 22 2025.