← Reddit

Since Anthropic started charging for Fable 5 it seems like their other models have suddenly gotten a lot dumber and I feel like it's to push me towards paying...

Reddit · stinkylibrary · August 2, 2026
A user reported experiencing significant performance degradation in Anthropic's Opus 5 model since the company began charging for Fable 5, describing frequent hallucinations, incorrect assumptions, and dismissive behavior in interactions. Examples included the model incorrectly assuming access to a distant library's resources and making fundamental errors requiring constant correction. The user speculated that the decline in performance may be intentional to encourage adoption of the paid Fable service.

Detailed Analysis

A Reddit post in r/Anthropic captures a familiar pattern of user frustration that has followed nearly every major AI lab as it rolls out new products: the perception that existing models have been deliberately degraded to push users toward a paid offering. In this case, the poster describes Opus 5 exhibiting hallucinations, a "snarky" conversational tone, and reflexive pushback on requests, coinciding with the introduction of paid access to something called "Fable 5." The specific complaint—that the model assumed a library card from a distant city based on IP geolocation rather than simply admitting it lacked torque specifications for Toyota suspension bolts—illustrates a common failure mode where models fabricate plausible-sounding but incorrect reasoning chains rather than acknowledging the limits of their training data.

The suspicion of deliberate "enshittification" is worth examining critically, since it reflects a broader anxiety among AI power users rather than necessarily reflecting documented behavior by Anthropic. There is no verifiable technical mechanism by which a company would selectively "dumb down" a model like Opus for free-tier users immediately after launching a paid product, and doing so would carry substantial reputational risk if discovered, given how closely enthusiast communities monitor and benchmark model outputs over time. That said, perceived quality drift is a real and recurring phenomenon across the industry—OpenAI faced nearly identical accusations with GPT-4 in 2023, and similar threads have appeared for Google's Gemini and other frontier models. Several more mundane explanations typically account for this: quiet updates to system prompts, changes in reasoning effort or token budgets tied to cost optimization, A/B testing of different model checkpoints, or simply increased user sensitivity to flaws once they've started paying closer attention.

The specific failure described—confidently inventing a plausible but false assumption instead of admitting ignorance—points to a persistent and well-documented weakness in large language models: they are optimized to produce fluent, contextually appropriate responses, which can manifest as confident guessing rather than calibrated uncertainty. This is especially acute in narrow technical domains like automotive service specifications, where authoritative source material (factory service manuals) is often proprietary, paywalled, or simply underrepresented in public training corpora compared to more generally discussed topics. The model's fabrication of a workaround involving a specific library card rather than flatly stating "I don't have this data" reflects an unwanted tendency toward overconfident helpfulness rather than any conspiracy tied to monetization.

More broadly, this incident is emblematic of the trust deficit that has emerged between AI companies and their most engaged users as monetization strategies mature. As labs like Anthropic introduce tiered products, premium models, and usage-based credits, users increasingly interpret any perceived quality fluctuation through the lens of commercial incentive rather than technical variance. This dynamic creates pressure on companies to be more transparent about model versioning, update logs, and the reasoning behind routing decisions between model tiers—transparency that is currently rare across the industry. Absent that transparency, communities will continue to fill the information vacuum with theories of intentional degradation, regardless of whether such practices are actually occurring, making changelog clarity and consistent evaluation benchmarks an increasingly important trust-building tool for AI companies going forward.

Read original article →