Skip to main content
Kaino.dev
Discover
Evals
News
Academics
Insights
Kaino.dev

Discover, evaluate, and compare AI tools, models, and agents.

Explore

  • Discover
  • Evaluations
  • News
  • Academics
  • Insights

Community

  • Twitter
  • YouTube
  • Instagram
Privacy PolicyTerms of Service

© 2026 Kaino.dev. All rights reserved.

Version 1.1.0
Claude Fable · Evaluations · Kaino
Claude Fable logo
model evaluation

Claude Fable

Anthropic

Claude Fable 5 is an Anthropic Claude model described as a Mythos-level model for ambitious long-running work.

modelanthropicclaudemythos
83.5KAINO SCORERecommended
Evaluated Jul 31, 202610 reviews
Website Docs GitHub

Scorecard

PricingMultimodalCostDev expTechnicalSpeedCodingReasoningRiskAdoption
  • Technical capability94
  • Coding & agentic94
  • Reasoning & knowledge89
  • Pricing clarity87
  • Developer experience87
  • Multimodal & I/O84
  • Adoption signal80
  • Speed & availability78
  • Risk & evidence76
  • Cost effectiveness66

Kainotomic evaluation

The supplied evidence identifies the target as Claude Fable 5, rather than the catalog label “Claude Fable.” It has unusually strong capability evidence: Anthropic reports 95.0% on SWE-bench Verified, DeepSWE lists 70%±4% (third behind Claude Opus 5 and GPT-5.6 Sol), and Terminal-Bench lists Claude Code with Fable 5 first at 83.8%±1.2%. These results support a coding score just below GPT-5.5/GPT-5.6 Sol’s 96/95 anchors, while its 1M context and 128k maximum output support high technical capability. Arena’s rank-one 1507±6 is supportive preference evidence, not independent proof of reasoning quality. Fable 5 compares favorably with Anthropic’s published Opus 4.8 anchor on context, documented agent performance, and apparent throughput; however, its coding score remains slightly lower because the strongest SWE-bench result is provider-reported and DeepSWE trails the stated leaders. Its $10/MTok input and $50/MTok output price is materially less costly than a typical premium frontier tier, but output pricing remains expensive relative to efficient models such as Gemini Flash or Kimi. Artificial Analysis reports 62 output tokens/s, a solid but not exceptional speed signal. Developer experience is strong through the Claude API, explicit model ID, caching prices, and long-context limits. Pricing documentation is detailed, but availability is materially unclear: the supplied product-page excerpt says unavailable while API documentation says general availability began June 9, 2026. Multimodal capability is not sufficiently specified in the provided excerpts, so it should not receive frontier-leading credit. The evaluation should remain conditional on confirming that the official pages and current API access are live.

Strengths

  • 95.0% SWE-bench Verified reported in Anthropic’s system card.
  • Third on DeepSWE v1.1 at 70%±4%; Terminal-Bench 2.1 lists a first-place Claude Code configuration at 83.8%±1.2%.
  • 1M-token context, up to 128k output tokens, explicit API identifier, and detailed cache pricing.
  • Arena Text rank 1 provides a strong public preference signal.

Caveats

  • SWE-bench result is vendor-reported; DeepSWE and Terminal-Bench are stronger independent corroboration for coding-agent performance.
  • Arena ranking is pairwise preference evidence rather than a controlled measure of technical capability.
  • Provided official excerpts do not establish the breadth or quality of image, audio, video, or tool I/O support.
  • $50/MTok output pricing limits cost effectiveness for output-heavy workloads.