Model Extraction Becomes a Policed Harm Category
Frontier labs increasingly instrument their APIs to detect and publicly attribute large-scale distillation by rival labs, so expect recurring threat reports naming corporate extractors, account-level enforcement against them, and pressure to treat paid API access as a controlled export rather than an open commercial product.
weakening · confidence 44 · Emerging (watchlist) · tracking since September 11, 2026 · updated September 14, 2026
Score history
Daily conviction score, 0 to 100. Higher means the thesis is more strongly corroborated.
Now 44 · -2 since Sep 13 · ranged 44 to 46
Showing the last few days. Unlock full score history.
Why the conviction moved
- Sep 12Strengthened +8
Anthropic publicly attributed 5,380 fraudulent Moonshot-linked accounts relaying roughly 300,000 requests in ten days, plus more than 12 million distillation attempts it assigns to DeepSeek over 14 days in July. This is the first named, quantified corporate attribution of large-scale extraction by rival labs, moving the practice from suspicion into an enforcement and policy category and strengthening the case for treating paid API access as controlled rather than openly commercial.
- Sep 11Strengthened
Anthropic's September threat report quantifies an alleged Alibaba distillation campaign at 151 million Claude exchanges and elevates model extraction to a named harm category alongside cyber operations and weapons research, with DeepSeek, Moonshot AI and MiniMax cited in smaller efforts.
Source trail
Supporting · September 12, 2026
Anthropic Says DeepSeek and Moonshot Routed Customer Prompts to Claude Through Fake Accounts
Anthropic publicly attributed 5,380 fraudulent Moonshot-linked accounts relaying roughly 300,000 requests in ten days, plus more than 12 million distillation attempts it assigns to DeepSeek over 14 days in July. This is the first named, quantified corporate attribution of large-scale extraction by rival labs, moving the practice from suspicion into an enforcement and policy category and strengthening the case for treating paid API access as controlled rather than openly commercial.
AnthropicSupporting · September 11, 2026
Anthropic Ties 151 Million Claude Exchanges to an Alleged Alibaba Distillation Campaign
Anthropic's September threat report quantifies an alleged Alibaba distillation campaign at 151 million Claude exchanges and elevates model extraction to a named harm category alongside cyber operations and weapons research, with DeepSeek, Moonshot AI and MiniMax cited in smaller efforts.
Anthropic News
Unlock full source trail, score history, and daily updates.
Unlock TrendsAffected regions & assets
Townsquare
Argue the thesis in Townsquare.