@tante The hard fact is that the statistical destillation process what some are calling training of LLMs is at the least partially transformative when it comes to copyright. There is precedent for that with concordances and the likes. The equally inconvenient fact is also that some of the inputs survive that process sufficiently unscathed that the result is blatant copyright infringement.