Claude Fable 5.1vsLLaMA 3.1

Claude Fable 5.1
LLaMA 3.1
Specifications
Context window
1M
API pricing
Input price
$10.00
Output price
$50.00
Cached input price
$0.25
Cheapest input
$10.00Amazon Bedrock$0.02DeepInfra
Cheapest output
$50.00Amazon Bedrock$0.04DeepInfra
Benchmarks
BullshitBench v2
77%14%
Benchmarks
SWE-Bench Pro
81.2%
SWE-Bench Multilingual
89.1%
SWE-Bench Multimodal
54.7%
Next.js Evals
97%
Terminal-Bench 4.0
57.88%
Terminal-Bench-Science 0.1
52.6%
Humanity's Last Exam · no tools
60.9%
Humanity's Last Exam · with tools
65%
ARC-AGI-2
90%
AutomationBench
31.4%
HealthBench Professional
62.1%
GDPval-AA v2
1853
AA-Briefcase
1694
threejseval
2037
Overview
CompanyAnthropicMeta
Release dateSep 1 2026Jul 23 2024
AccessProprietaryOpen Weight

Other comparisons

Claude Fable 5.1vsGPT-6 AstraLLaMA 3.1vsGPT-6 AstraClaude Fable 5.1vsGemini 3.8 FlashLLaMA 3.1vsGemini 3.8 FlashClaude Fable 5.1vsGrok 4.6LLaMA 3.1vsGrok 4.6Claude Fable 5.1vsDeepSeek-V4.1-FlashLLaMA 3.1vsDeepSeek-V4.1-FlashClaude Fable 5.1vsMistral Medium 3.5LLaMA 3.1vsMistral Medium 3.5Claude Fable 5.1vsKimi K3LLaMA 3.1vsKimi K3

Frequently asked questions

Claude Fable 5.1 leads LLaMA 3.1 on 1 of the 1 benchmark they both report (BullshitBench v2). Only Claude Fable 5.1 has a verified first-party API price: $10.00 per million input tokens and $50.00 per million output tokens. No pay-as-you-go API rate is tracked for LLaMA 3.1. LLaMA 3.1 shipped 770 days before Claude Fable 5.1, so benchmark comparisons should account for the intervening progress.

Claude Fable 5.1 is proprietary, while LLaMA 3.1 is open weight.

On BullshitBench v2, Claude Fable 5.1 leads at 77% vs LLaMA 3.1 at 14%.