12 tracked signals on distillation.
Open letters about AI development
Jensen Huang · Simon Willison · Aug 02, 2026
Competing open letters expose industry fracture over open-weight AI models and development pacing
“We request that the U.S. government support an international effort to develop the technical and governance tools needed to deliberately pace the frontier of automated AI development.”
Disrupting a coordinated model-distillation campaign
OpenAI · OpenAI Blog · Sep 30, 2026
OpenAI disrupted a coordinated adversarial campaign to extract protected model reasoning via distillation
DeepSeek Just Made Closed AI Look Ridiculous
Two Minute Papers · Aug 19, 2026
DeepSeek 4 Pro achieves near-frontier quality with MIT-licensed open weights, pressuring closed AI labs.
“DeepSeek has MIT licensed open weights. Anyone can run the exact same model at their own price.”
Open models recap: more on Kimi K3, Qwen 3.8, Xi's WAIC speech, distillation, the open-closed gap, and what's next
Nathan Lambert · Interconnects · Jul 22, 2026
China commits to AI openness as Xi's WAIC speech and Qwen's open-weight pivot signal strategic shift
“Xi gave his speech where he directly committed to openness and open source as a strategy.”
Who’s Afraid of Chinese Models?
Simon Willison · Jul 20, 2026
Ben Thompson proposes US law legalizing AI training data collection and distillation to compete with China
“The U.S. should pass a law that (1) makes explicit that collecting data for training models is fair use, and (2) bars terms of service that forbid distillation, for U.S. companies at a minimum.”
Frontier post-training recipe review with Finbarr Timbers
Nathan Lambert · Interconnects · Jun 16, 2026
2026 frontier post-training has shifted to Multi-Teacher On-Policy Distillation (MOPD), merging many specialist models into one.
“The shape of a post-training recipe has changed more in the last year than in the prior three.”
Why Distillation Might Be Impossible to Stop
LangChain · Jul 21, 2026
Distillation is structurally unpreventable because all models converge toward a shared platonic representation of intelligence
“distillation is effectively unpreventable. Because once you've found the shape of intelligence, you know what it looks like.”
Deploy. Observe. Learn. Reinforcement learning for production agents | BRK231
Microsoft Developer (Build) · Jun 03, 2026
Microsoft Foundry adds reinforcement learning and a low-level training API to turn production agents into cheaper, faster models.
“Think of it as PyTorch as a service.”
Generative Video at the Speed of Light — Keegan McCallum, uRun
AI Engineer · Aug 18, 2026
uRun's Helios model generates video at 1/100th the cost of frontier models with comparable quality
“at least 40 models with real-time capabilities and long horizon generation capabilities released this year”
Hugging Face Journal Club: Direct On-Policy Distillation
Hugging Face · Aug 11, 2026
On-policy distillation uses RL policy shift signals to train student models without full RL cost
“the student might be stronger than the teacher in this case”
LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation
Hugging Face · Hugging Face Blog · Aug 19, 2026
Hugging Face released LFM2.5 Q4_0 checkpoints via quantization-aware distillation
Bringing Continual Learning into Enterprises — Samuel Denton, Applied Compute
AI Engineer · Aug 12, 2026
Applied Compute is bringing continual learning to enterprises via a distillation spectrum from offline to online
“this is sort of the holy grail of continual learning where I have a model that's serving production traffic, it does a rollout, it creates a trace, we figure out how to learn from that trace, we update the model”