Charlie O'Neill, Co-Head of model training at Baseten, argues that distillation is relatively unimportant for competitive advantage, citing Sonnet 5's inferior performance compared to GLM 5.2 despite likely distillation from Mythos. He notes that distillation primarily serves as a warmstart for reinforcement learning, with Chinese labs achieving similar model quality using tokens from American models, albeit at a higher cost.