Comment Re:Pot meet kettle (Score 1) 75
I don't think training a model with copyrighted data is any kind of gray zone at all.
Running inference against such a model _might_ be infringing.
But I don't see the legal argument that model training would be.
Think of model training as lossy data compression. You provide an input, it produces a compressed artifact, and it can no longer reproduce the original when you uncompress it. It can produce _something_. It cannot reproduce the original.
On slashdot, we all agree that format shifting copyrighted works should be allowed and unrestricted, right?
If there is applicable case law here about how lossy a "reproduction" can be before copyright has something to say about it, that would be relevant. Do you know of any?