← Hacker News

Anthropic says Alibaba used 25k accounts to mine Claude

Hacker News · logickkk1 · June 27, 2026

Detailed Analysis

Anthropic has accused Alibaba, the Chinese e-commerce and technology conglomerate, of orchestrating a large-scale, coordinated effort to systematically extract data from Claude using approximately 25,000 separate accounts. The allegation represents one of the most significant publicly disclosed instances of alleged AI model scraping by a major technology company against a competitor, and raises serious questions about intellectual property, terms-of-service enforcement, and competitive conduct in the global AI industry. The sheer scale of the operation — spanning tens of thousands of accounts — suggests a deliberate, resourced effort rather than opportunistic individual use.

The practice commonly referred to as "model mining" or "model scraping" involves querying an AI system at high volume and in structured ways to extract its underlying behaviors, reasoning patterns, or response characteristics. This data can then theoretically be used to train competing models, to benchmark capabilities, or to replicate proprietary fine-tuning. For Anthropic, whose Claude models represent core commercial and research assets developed at considerable expense, unauthorized systematic extraction of this kind would constitute a direct threat to its competitive position and potentially to the safety properties it has embedded in the model through its Constitutional AI methodology.

The accusation carries broader geopolitical dimensions given the identity of the alleged actor. Alibaba operates Tongyi Qianwen and other large language model products that compete directly in the global AI market, and Chinese technology firms have faced heightened scrutiny from Western governments and companies over data practices and competitive behavior. Anthropic's decision to publicly attribute the activity to Alibaba — rather than resolving it quietly — signals an intent to draw regulatory and public attention to what it characterizes as systematic abuse of its platform, a posture that may also serve as a deterrent to similar actors.

This episode fits within a broader and escalating pattern of tension between AI developers over data, model capabilities, and competitive intelligence. OpenAI has similarly taken legal and technical action against parties it accused of scraping its systems, and AI companies across the industry have invested heavily in detection systems designed to identify coordinated inauthentic usage at scale. The use of 25,000 accounts specifically points to sophisticated infrastructure designed to evade rate limits and detection thresholds that would flag individual high-volume users, underscoring how technically advanced model-scraping operations have become. As AI models become more economically valuable, enforcement of terms of service and the legal frameworks surrounding model-derived data are expected to become increasingly contested terrain in courts and regulatory bodies worldwide.

Read original article →