RL economics, morally charged terms, and "distillation"

240 · Addxorrol Blogspot · June 15, 2026, 9:10 a.m.
Summary
This article discusses the economics of training large language models (LLMs) using reinforcement learning (RL) and critiques the moral framing used by model providers regarding copyright and model training. The author argues against the use of the term 'distillation' in this context, suggesting it carries misleading moral implications that serve the interests of major players in the LLM space.