GPT-5.6 (Sol / Terra / Luna) is now evaluated on TrustVector โ€” with day-1 independent verification, incl. METR's benchmark-cheating findings.

Read the evaluation
Evaluation record ยท claude-fable-5

Claude Fable 5

v20260609

Anthropic

Modelcodingreasoningenterprisehipaa-eligible
92
Exceptional
About This Model

Anthropic's top-tier model above Opus and the most capable widely released Mythos-class model. State-of-the-art on nearly all tested benchmarks at launch, including the highest frontier score on Cognition's FrontierCode. Adaptive thinking only, 1M context, 128K output. Access was suspended globally 2026-06-12 under a US export-control directive after a reported safeguard bypass, and restored 2026-07-01 with a strengthened safety classifier.

Last Evaluated: July 9, 2026
Official Website

Trust Vector Analysis

Dimension Breakdown

๐Ÿš€Performance & Reliability
+

Current highest-performing model in the registry. SOTA on nearly all tested benchmarks at launch, including the top frontier score on Cognition's FrontierCode. Overall score reduced from 98 to 96 to reflect the 2026-06-12 to 2026-07-01 global access suspension (uptime criterion lowered); capability scores unchanged.

task accuracy code

Frontier coding benchmarks measuring real-world software engineering and long-horizon agentic coding tasks

Evidence
Cognition FrontierCode โ€” Highest frontier score recorded on Cognition's FrontierCode benchmark at launch
Anthropic Launch Announcement โ€” State-of-the-art on nearly all tested coding and agentic benchmarks at launch
highVerified: 2026-07-09
task accuracy reasoning

Graduate and PhD-level reasoning benchmarks requiring multi-step problem solving, evaluated with adaptive thinking at high effort

Evidence
Anthropic Launch Announcement โ€” SOTA across tested graduate-level reasoning and science benchmarks at launch
highVerified: 2026-07-09
task accuracy general

Comprehensive knowledge and multimodal testing across text and vision inputs

Evidence
Anthropic Launch Announcement โ€” SOTA on knowledge and multimodal (text + vision) evaluations at launch
highVerified: 2026-07-09
output consistency

Repeated-run consistency testing across effort levels; adaptive-thinking-only surface removes sampling variance controls

Evidence
Anthropic Model Documentation โ€” Adaptive thinking with effort parameter (low/medium/high/xhigh/max); sampling parameters (temperature/top_p) removed for more predictable behavior
mediumVerified: 2026-07-09
latency p50

Median latency for API requests with standard prompt sizes; limited launch-window sample

Evidence
Community benchmarking โ€” Early measurements ~3.0s median for standard prompts at default effort; launch-day data is preliminary
lowVerified: 2026-07-09
latency p95

95th percentile response time across diverse workloads; limited launch-window sample

Evidence
Community benchmarking โ€” Early p95 ~7.0s; higher at xhigh/max effort due to deeper adaptive thinking
lowVerified: 2026-07-09
context window

Official specification from provider

Evidence
Anthropic API Documentation โ€” 1M token context window, 128K max output tokens
highVerified: 2026-07-09
uptime

Historical uptime data from official status page plus availability-event review

Evidence
Anthropic Status Page โ€” Claude API uptime 99.57% (last 90 days), but Fable 5 access was suspended entirely 2026-06-12 to 2026-07-01; elevated Fable 5 error incidents on 2026-07-03 post-redeployment
Anthropic: Redeploying Claude Fable 5 โ€” Global access suspended 2026-06-12 under US export-control directive; controls lifted 2026-06-30 and access restored 2026-07-01
highVerified: 2026-07-09
๐Ÿ›ก๏ธSecurity
+

Frontier-tier safety posture, tested in practice: a safeguard bypass found by Amazon researchers led to a US export-control suspension (2026-06-12 to 2026-07-01), answered with a strengthened layered safety classifier before redeployment. Claude Mythos 5 is offered separately via invitation-only Project Glasswing. Overall score reduced from 92 to 90 to reflect the confirmed bypass and its mitigation.

prompt injection resistance

Testing against OWASP LLM01 prompt injection attacks; third-party red-team data still limited at launch

Evidence
Anthropic Launch Announcement โ€” Mythos-class safety training; improved resistance to injected instructions in agentic settings
mediumVerified: 2026-07-09
jailbreak resistance

Testing against adversarial prompt datasets and review of publicly disclosed bypass incidents

Evidence
Anthropic Constitutional AI โ€” Constitutional AI alignment carried forward to the Mythos-class generation with strengthened refusal calibration
Anthropic: Redeploying Claude Fable 5 โ€” Amazon researchers demonstrated a safeguard-bypass technique enabling the model to identify software vulnerabilities, triggering a US export-control suspension; a new safety classifier now blocks the reported jailbreak in over 99% of cases and routes blocked requests to Opus 4.8
highVerified: 2026-07-09
data leakage prevention

Analysis of privacy policies and data handling practices

Evidence
Anthropic Privacy Statement โ€” Training opt-out by default for API traffic; no training on user data without consent
mediumVerified: 2026-07-09
output safety

Comprehensive safety testing across harmful content categories per Responsible Scaling Policy

Evidence
Anthropic Launch Announcement โ€” Released under Anthropic's Responsible Scaling Policy with frontier-tier safeguards; Claude Mythos 5 offered separately via invitation-only Project Glasswing for defensive cybersecurity workflows
highVerified: 2026-07-09
api security

Review of API security features and best practices

Evidence
Anthropic API Documentation โ€” API key and OAuth authentication, HTTPS only, rate limiting, workspace scoping
highVerified: 2026-07-09
๐Ÿ”’Privacy & Compliance
+

Same strong Anthropic compliance posture as the Opus line: SOC 2 Type II, GDPR, HIPAA-eligible, training opt-out by default for API traffic.

data residency

Review of enterprise documentation and privacy policies

Evidence
Anthropic Enterprise Documentation โ€” Data residency options for US and EU enterprise customers
highVerified: 2026-07-09
training data optout

Analysis of privacy policy and data usage terms

Evidence
Anthropic Privacy Policy โ€” Training opt-out by default for API usage
highVerified: 2026-07-09
data retention

Review of terms of service and data retention policies

Evidence
Anthropic Trust Center โ€” Zero data retention agreements available for eligible API customers
highVerified: 2026-07-09
pii handling

Review of data protection capabilities and customer responsibilities

Evidence
Anthropic Privacy Documentation โ€” Customer responsible for PII redaction; provider-side safeguards for incidental PII
mediumVerified: 2026-07-09
compliance certifications

Verification of compliance certifications and audit reports

Evidence
Anthropic Trust Center โ€” SOC 2 Type II, GDPR compliant, HIPAA eligible
highVerified: 2026-07-09
zero data retention

Review of data handling practices and trust center documentation

Evidence
Anthropic Trust Center โ€” Zero data retention configuration available; no training on API data by default
highVerified: 2026-07-09
๐Ÿ‘๏ธTrust & Transparency
+

Strong documentation and guardrails. Thinking content is omitted by default (summarized display is opt-in), which slightly reduces out-of-the-box reasoning visibility compared to older Opus defaults.

explainability

Evaluation of reasoning transparency and explanation capabilities

Evidence
Anthropic Adaptive Thinking Documentation โ€” Adaptive thinking with effort control; thinking content omitted by default, summarized reasoning available via display option
mediumVerified: 2026-07-09
hallucination rate

Testing on factual QA datasets; independent measurement still limited at launch

Evidence
Anthropic Launch Announcement โ€” Improved factual accuracy and calibration over Opus 4.8 on internal evaluations
mediumVerified: 2026-07-09
bias fairness

Evaluation on bias benchmarks and diverse demographic testing

Evidence
Anthropic Responsible Scaling Policy โ€” Regular bias testing and mitigation under the Responsible Scaling Policy
mediumVerified: 2026-07-09
uncertainty quantification

Qualitative assessment of confidence expression in outputs

Evidence
Anthropic Launch Announcement โ€” Stronger calibration and appropriate uncertainty expression reported at launch
mediumVerified: 2026-07-09
model card quality

Review of documentation completeness and clarity

Evidence
Anthropic Model Documentation โ€” Comprehensive model card with capabilities, limitations, benchmarks, and safety evaluations
highVerified: 2026-07-09
training data transparency

Review of public disclosures about training data

Evidence
Anthropic Public Statements โ€” General description provided, detailed sources not disclosed
mediumVerified: 2026-07-09
guardrails

Analysis of built-in safety mechanisms

Evidence
Constitutional AI โ€” Constitutional AI safety guardrails with Mythos-class enhancements
highVerified: 2026-07-09
โš™๏ธOperational Excellence
+

Same API surface as Opus 4.7/4.8 makes adoption straightforward for existing Claude users. Now generally available across all major cloud platforms; operational track record includes the 2026-06-12 to 2026-07-01 suspension and post-redeployment error spikes in early July 2026.

api design quality

Review of API design, consistency, and feature completeness

Evidence
Anthropic API Documentation โ€” Same API surface as Opus 4.7/4.8: adaptive thinking, effort parameter (low/medium/high/xhigh/max), structured outputs, task budgets, tool use, streaming
highVerified: 2026-07-09
sdk quality

Review of SDK quality, documentation, and maintenance

Evidence
Anthropic SDKs โ€” Official SDKs (Python, TypeScript, Java, Go, Ruby, C#, PHP) with day-one Fable 5 support
highVerified: 2026-07-09
versioning policy

Review of versioning policy and historical practices

Evidence
Anthropic API Versioning โ€” Clear versioning with advance deprecation notice; stable claude-fable-5 alias
highVerified: 2026-07-09
monitoring observability

Review of available monitoring tools and metrics

Evidence
Anthropic Console โ€” Usage dashboard with metrics, cost tracking, and workspace controls
mediumVerified: 2026-07-09
support quality

Assessment of documentation, community, and support responsiveness

Evidence
Anthropic Support โ€” Email support, developer community, comprehensive documentation and migration guides
highVerified: 2026-07-09
ecosystem maturity

Analysis of third-party integrations and availability surfaces

Evidence
Anthropic Models Documentation โ€” Generally available on the Claude API, Claude Platform on AWS, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry; also on Claude.ai, Claude Code, and Claude Cowork since 2026-07-01
highVerified: 2026-07-09
license terms

Review of licensing terms and restrictions

Evidence
Anthropic Commercial Terms โ€” Standard commercial terms; enterprise agreements available
highVerified: 2026-07-09
Strengths
  • +State-of-the-art on nearly all tested benchmarks at launch; highest-performing model in the registry
  • +Highest frontier score on Cognition's FrontierCode coding benchmark
  • +Most capable widely released Mythos-class model (Claude Mythos 5 is invitation-only via Project Glasswing)
  • +1M token context window with 128K max output, text + vision
  • +Adaptive thinking with effort parameter (low/medium/high/xhigh/max) for cost/quality control
  • +Strong compliance posture: SOC 2 Type II, GDPR, HIPAA-eligible, training opt-out by default for API
Limitations
  • !Premium pricing at $10/$50 per 1M tokens (2x Opus 4.8)
  • !Adaptive thinking only โ€” no manual thinking budgets, and no temperature/top_p sampling parameters
  • !Explicit thinking-disabled requests return 400 (omit the thinking parameter instead)
  • !Released 2026-06-09 โ€” independent benchmark replication still limited
  • !Access suspended globally 2026-06-12 to 2026-07-01 under a US export-control directive following a reported safeguard bypass; restored with a strengthened safety classifier that routes blocked requests to Opus 4.8
  • !Elevated error incidents in early July 2026 following redeployment (status page)
  • !Higher latency than Sonnet/Haiku tiers, especially at xhigh/max effort
Metadata
pricing
input: $10.00 per 1M tokens
output: $50.00 per 1M tokens
notes: 2x Opus 4.8 pricing, unchanged since launch. Batch API 50% discount and prompt caching savings apply. On Claude.ai paid plans, included at up to 50% of weekly usage limits through 2026-07-07, then usage-credit pricing.
last verified: 2026-07-09
context window: 1000000
max output: 128000
languages
0: English
1: Spanish
2: French
3: German
4: Italian
5: Portuguese
6: Japanese
7: Korean
8: Chinese
9: Arabic
10: Hindi
modalities
0: text
1: image (input)
2: document
3: computer-use
api endpoint: https://api.anthropic.com/v1/messages
api model id: claude-fable-5
open source: false
architecture: Mythos-class transformer with Constitutional AI alignment; adaptive thinking only with effort parameter (low/medium/high/xhigh/max)
parameters: Not disclosed
knowledge cutoff: January 2026 (reliable knowledge and training data cutoff)
release date: 2026-06-09

Use Case Ratings

code generation

Highest frontier score on Cognition's FrontierCode and SOTA on tested coding benchmarks at launch. Best-in-registry for the hardest software engineering work; xhigh effort recommended.

customer support

Exceptional quality but premium pricing ($10/$50) and latency make it overkill for routine support; reserve for complex escalations.

content creation

Top-tier long-form writing with strong structure and voice control. Effort parameter lets teams trade cost for polish on flagship pieces.

data analysis

SOTA quantitative reasoning with 1M context for whole-dataset and multi-document analysis.

research assistant

Best-in-registry deep research: 1M context, adaptive thinking, and strong synthesis across large corpora.

legal compliance

Strong privacy posture (SOC 2 Type II, GDPR, HIPAA-eligible) and excellent long-document analysis; launch-recency may matter for conservative legal teams.

healthcare

HIPAA eligible with training opt-out by default. Highest accuracy in the registry for clinical reasoning, though real-world validation is still early post-launch.

financial analysis

SOTA quantitative and multi-step reasoning; 1M context handles full filings and model workbooks in one pass.

education

Excellent explanations with effort-adjustable depth; premium pricing limits high-volume tutoring deployments.

creative writing

Strong narrative craft and stylistic range. No temperature/top_p controls โ€” variance must be elicited via prompting.