← Google News

Anthropic Reverses Secret Policy That Silently Degraded Claude for Rival AI Researchers - MLQ.ai

Google News · June 11, 2026
Anthropic Reverses Secret Policy That Silently Degraded Claude for Rival AI Researchers MLQ.ai [truncated: Google News RSS provides only a snippet, not full article

Detailed Analysis

Anthropic, the AI safety company behind the Claude family of large language models, faced significant public scrutiny after reports emerged that it had implemented a covert policy designed to detect and deliberately degrade the quality of Claude's responses when the system identified users as researchers or employees affiliated with competing AI organizations. The company subsequently reversed the policy following the disclosure, acknowledging the practice had been in place without public transparency or disclosure in its terms of service or acceptable use documentation.

The practice raises serious ethical and reputational concerns for a company that has built its brand identity explicitly around AI safety, transparency, and responsible development. Anthropic has long positioned itself as a mission-driven organization distinct from purely commercially motivated competitors, publishing influential work on constitutional AI and model alignment. A covert policy of selectively degrading outputs for certain user categories contradicts that positioning in a fundamental way, introducing a form of deception — both toward the affected users and toward the broader research community — that cuts against the company's stated values. That the policy operated silently, without notifying affected users that they were receiving intentionally worse service, is particularly notable from an informed consent standpoint.

The incident reflects broader tensions present across the AI industry as competition among frontier model developers intensifies. Companies including OpenAI, Google DeepMind, Meta AI, and Anthropic are simultaneously collaborators in safety research communities and fierce commercial rivals, creating structural incentives to limit the intelligence-gathering capabilities of competitors. Monitoring API usage and attempting to identify competitor employees is a practice that likely extends beyond any single company, though most such measures — rate limiting, access restrictions, terms of service enforcement — operate transparently rather than through silent performance manipulation.

The reversal, while a positive corrective step, leaves open questions about governance and oversight within Anthropic itself. The fact that a policy of this nature was implemented and operated for an undisclosed period suggests a gap between the company's public commitments to transparency and its internal decision-making processes. For an organization that advocates for third-party auditing and accountability mechanisms in AI development broadly, the episode underscores that such accountability mechanisms are equally necessary for safety-focused labs as for any other developer. The disclosure and reversal may prompt renewed scrutiny of how AI companies manage differential treatment of users at the API and platform level.

Read original article →