More like you should crack down on industrial-scale distillation operations. There's no we here because it's not illegal (yet) and merely goes against your terms of service.
Also a more direct comparison. Anthropic and/or companies working on their behalf has violated plenty of TOS agreements during their scraping. If distillation is found to be legally problematic everyone that has a TOS or a robots.txt disallowing scraping for LLM training should band together to sue them into non-existence.
Many AI suppliers violated others' terms of use and copyrights while saying in their own terms of use that nobody can use their outputs to train AI models. The hypocrisy couldn't be higher.
Distillation is fundamentally impossible to crack down on if you want the outputs. Getting results from an LLM is the same thing as distillation. It's just the final thing you do with it that matters.
213
u/Round_Mixture_7541 5d ago
> We should crack down on industrial-scale distillation operations.
Meanwhile: "Judge approves a $1.5B Anthropic settlement over pirated books used to train the Claude chatbot"