Caught Distilling from Claude?
Quick Overview
Anthropic publicly accused three AI labs—DeepSeek, Moonshot, and MiniMax—of orchestrating industrial-scale distillation attacks against their Claude models, involving over 16 million exchanges generated through approximately 24,000 fraudulent accounts to illicitly extract capabilities like agentic reasoning and coding.
Key Points: Anthropic identified industrial-scale distillation campaigns targeting their Claude models by three AI laboratories: DeepSeek, Moonshot, and MiniMax. The total illicit activity involved over 16 million exchanges generated through approximately 24,000 fraudulent accounts, violating terms of service. DeepSeek accounted for over 150,000 exchanges, targeting reasoning capabilities and using rubric-based grading tasks to distill Claude's function as a reward model for reinforcement learning. Moonshot AI employed hundreds of fraudulent accounts across multiple access pathways, targeting agentic reasoning, coding, tool use, and computer vision, totaling over 3.4 million exchanges. MiniMax conducted the largest operation with over 13 million exchanges, focusing on agentic coding and tool use/orchestration. The methodology involved prompting Claude to articulate internal reasoning to generate chain-of-thought training data at scale, which is key to illicit knowledge extraction. The speaker notes that Anthropic was sued last year for using over 7 million pirated books to train Claude, resulting in a $1.5 billion settlement, highlighting ongoing data sourcing controversies in the industry.
Context: The video discusses Anthropic's public announcement regarding large-scale attempts by competing AI labs to extract proprietary knowledge from their Claude language models using a technique called 'distillation.' The speaker references Anthropic's official blog post detailing how DeepSeek, Moonshot, and MiniMax allegedly used tens of thousands of fake accounts to query Claude extensively for its advanced capabilities. This context is further framed by referencing a major lawsuit against Anthropic itself for using copyrighted material in its own training data, suggesting a complex ethical landscape around model training and knowledge extraction in the AI industry.