Nothing against what I said in there. Distillation is far from enough to get to the frontier. Its at best a ramp (the most efficient one used by everyone ) that shortcut and saves millions of rl runs before a model moves.
the initial post infers that distillation is all you need. it is not. in 2026 you need large scale distributed inference of gigantic models, rl envs and millions of dollars runs to get to something decent. if you think glm just has to sft on traces of claude to edge Mythos on some cyber benchmark you are fooled.
I said there is no evidence distillation is all you need as per the initial comment "The fact that AI models can be so easily distilled and replicated is such a stroke of luck." it is far from trivial to reproduce the capabilities of anthropic models for instance