White House calls Chinese AI lab a copycat, but experts say Kimi K3 can't just be a Claude clone
Curated by the Inblix editorial team
The White House is lobbing a serious accusation at one of China’s most prominent AI labs. Science advisor Michael Kratsios claims Moonshot, the company behind the massive open-weight model Kimi K3, built its creation by illicitly copying Anthropic’s Fable LLM using smuggled Nvidia chips. “Large-scale, covert industrial distillation aimed at stealing proprietary U.S. technology and undermining American research is unacceptable,” Kratsios wrote, echoing Treasury Secretary Scott Bessent’s claim that they’re “finding watermarks” of American models on Chinese ones. Moonshot hasn’t commented, and the government hasn’t provided any evidence for the watermark claim.
But the technical community isn’t buying the simple copycat narrative. Researchers point to a glaring timeline problem. Anthropic’s Fable only became publicly available on July 1st, and Kimi K3 appeared just two weeks later. “You can’t distill that much data, train a model, and release it in two weeks,” Braden Hancock of the Laude Institute and Snorkel AI told TechCrunch. The sheer logistical impossibility of that schedule is making experts skeptical that distillation alone explains the model’s advanced capabilities.
The debate hinges on how modern AI models actually learn. Distillation is a known technique where a lab systematically queries a target model to extract its problem-solving methods, often through supervised fine-tuning (SFT). But Nathan Lambert, a researcher at the Allen Institute for AI, argues that SFT’s importance is fading as models reach the frontier and shift to reinforcement learning. To truly copy Fable-level reasoning, a lab would need to run millions of expensive, slow API queries through a reinforcement learning agent—a process that would be both a financial and time bottleneck. “If it were the case, everyone would be easily able to catch up,” Lambert said on a podcast, noting we haven’t seen that happen.
The reality is probably messier than a simple heist. Anthropic has previously accused Moonshot, DeepSeek, and MiniMax of systematic distillation earlier this year, discovering millions of unusual queries. But distillation is an industry-wide gray area, not a uniquely Chinese sin. Elon Musk testified that his own team distilled OpenAI models to build Grok, calling it common practice. Hancock adds that American observers consistently underrate Chinese engineering talent, noting Moonshot’s founder is a CMU PhD. “If American models ground to a halt, I think China’s progress would slow, but would still continue,” he said. “They’re not just riding coattails here.” The chip-smuggling allegation adds another layer, but it doesn’t settle the core question of whether Kimi K3 is a copycat or evidence of genuine, independent technical momentum.
💡 Key Takeaways
- The White House claims Moonshot copied Anthropic's Fable to build Kimi K3, but the July 1st release of Fable and July 15th launch of Kimi makes simple distillation an impossible timeline, according to researchers.
- AI researcher Nathan Lambert argues supervised fine-tuning is becoming less useful for copying frontier models, and that replicating Fable's capabilities via reinforcement learning would be prohibitively expensive and slow.
- Distillation is an industry-wide gray area—Elon Musk admitted to using it on OpenAI models for Grok—making a purely national security framing a selective interpretation of a common practice.
- Chinese AI teams include top-tier researchers like Moonshot's CMU PhD founder, and experts say their progress would continue even without access to American models, though likely at a slower pace.
Keep reading: See related articles below for more coverage on this topic.
Get smarter about AI
The sharpest AI news, curated daily. Delivered free to your inbox.