WINTHROPICindependent · unaffiliated
batch 2026-09-04 · 11:30 UTC
Answer page · regenerated with every reading · 2026-09-04

Is Claude Fable getting worse?

As of 2026-09-04, fable-5-1 scored 12/12 on Winthropic's fixed 12-task battery, unchanged from the previous run.

Last 5 readings
Runfable-5-1
2026-09-0412/12
2026-09-0312/12
2026-09-02-
2026-09-01-
2026-08-31-

Pass count on the same 12-task battery, one reading per day. Each date links to the full run record with every failure's output. Full trace: /m/fable-5-1.

How this is measured

The same versioned battery runs against every tracked Claude model each morning under a frozen request shape; scoring is deterministic and no model judges another. Claude 5 models always sample, so a one-point move can be variance; a move that stays is drift. Methodology.

Cite this reading
Winthropic Index, "Is Claude Fable getting worse?", reading of 2026-09-04, https://winthropic.com/is-claude-fable-getting-worse
@misc{winthropic_is_claude_fable_getting_worse_20260904,
  title = {Is Claude Fable getting worse?},
  author = {Winthropic},
  year = {2026},
  note = {Reading of 2026-09-04. Independent daily evaluation; not affiliated with Anthropic.},
  url = {https://winthropic.com/is-claude-fable-getting-worse}
}

Winthropic is an independent measurement project and is not affiliated with, endorsed by, or connected to Anthropic. If you are paying for output you are not getting, Winthrop's Token Audit prices your actual spend from a usage export; the summary is free.