Skip to content

Debates

Is distilling other labs' models a legitimate strategy, and should labs prohibit it?

5 recorded positions from 4 people, first said Feb 17, 2025. They do not agree — the readings below are what each one actually argued.

Also on the record

Steeve Morin · Feb 24, 2025

Distillation of other labs' models is fair game

There is no free lunch — the frontier models were themselves trained on other people's data, as shown by an image model reproducing a Star Wars screenshot on request; you took it from the beginning too

53:09 Distillation is fair game since frontier models themselves trained on others data no free lunch

Mike Krieger · Mar 3, 2025

Distillation within a lab is valuable, but no nation should be able to distill models from another nation's labs

Internally, distilling your highest-end model into lower-latency, cheaper versions is useful; across nations there is national security value in being thoughtful as AI capabilities grow

30:25 Cross national distillation should be restricted though intra lab distillation is fine

Alex Atallah · Aug 10, 2026 · hedged

American labs could get pretty far by distilling Chinese open-weight models, and doing so is safer than feared because the outputs are inspectable

Most Chinese open-weight models permit distillation, letting you do RL on their outputs; and because you see the outputs during RL rollouts, you can catch misalignment with your model's voice or constitution

46:51 Distilling chinese open weights is viable and inspectable

Alex Atallah · Aug 10, 2026

Labs have a legitimate right to prohibit distillation in their terms of service and cut off competitors, and a market will split between permissive and restrictive providers

A company can cut off access to someone building a competitive model, though most frontier labs don't prohibit building small non-competitive task-specific models

49:05 Labs may legitimately prohibit distillation and the market splits

Jonathan Ross · Feb 17, 2025

China's edge is a greater willingness to use methods Western labs treat as off-limits — DeepSeek distilled the OpenAI model, which most model providers considered a red line

Many argue OpenAI scraped the internet so distillation is fair game, but the established model providers had treated distilling a rival's model as a line they wouldn't cross

53:51 Chinas willingness to cross distillation red lines gives it an edge

Your assistant can query this graph directly — 5 positions here, 19,646 across the corpus. Add 996.fm over MCP.