The "distillation" argument always seemed pretty weak to me. If one argues that training an AI on copyrighted content is merely analogous to reading a book (and therefore is fully transformative), then it stands to reason that training an AI on other AI outputs would also be fully transformative.
It seems to me you can't have it both ways; either training is a violation of copyright or it isn't, and there's no consistent argument where "distillation" is a violation but other things aren't.
It seems to me you can't have it both ways; either training is a violation of copyright or it isn't, and there's no consistent argument where "distillation" is a violation but other things aren't.
https://www.techopedia.com/trump-administration-openai-train...