Claude Fable 5.1vsGPT-4.1

Claude Fable 5.1
GPT-4.1
Specifications
Context window
1M
API pricing
Input price
$10.00$2.00
Output price
$50.00$8.00
Cached input price
$0.25$0.50
Cheapest input
$10.00Amazon Bedrock$2.00Azure
Cheapest output
$50.00Amazon Bedrock$8.00Azure
Benchmarks
BullshitBench v2
77%14%
Benchmarks
SWE-Bench Pro
81.2%
SWE-Bench Verified
54.6%
SWE-Bench Multilingual
89.1%
SWE-Bench Multimodal
54.7%
Next.js Evals
97%
Terminal-Bench 4.0
57.88%
Terminal-Bench-Science 0.1
52.6%
Humanity's Last Exam · no tools
60.9%
Humanity's Last Exam · with tools
65%
ARC-AGI-2
90%
AutomationBench
31.4%
HealthBench Professional
62.1%
GDPval-AA v2
1853
AA-Briefcase
1694
threejseval
2037
Overview
CompanyAnthropicOpenAI
Release dateSep 1 2026Apr 14 2025
AccessProprietaryProprietary

Other comparisons

Claude Fable 5.1vsGemini 3.8 FlashGPT-4.1vsGemini 3.8 FlashClaude Fable 5.1vsMuse Spark 1.3GPT-4.1vsMuse Spark 1.3Claude Fable 5.1vsGrok 4.6GPT-4.1vsGrok 4.6Claude Fable 5.1vsDeepSeek-V4.1-FlashGPT-4.1vsDeepSeek-V4.1-FlashClaude Fable 5.1vsMistral Medium 3.5GPT-4.1vsMistral Medium 3.5Claude Fable 5.1vsKimi K3GPT-4.1vsKimi K3

Frequently asked questions

Claude Fable 5.1 leads GPT-4.1 on 1 of the 1 benchmark they both report (BullshitBench v2). GPT-4.1 is cheaper on both input and output: $2.00 vs $10.00 per million input tokens, and $8.00 vs $50.00 per million output tokens. GPT-4.1 shipped 505 days before Claude Fable 5.1, so benchmark comparisons should account for the intervening progress.

Published specifications for these two models are limited — see each model page for the latest details.

On BullshitBench v2, Claude Fable 5.1 leads at 77% vs GPT-4.1 at 14%.