We got asked this question two weeks ago when we conducted a workshop on effective and responsible use of AI tools at a local conference. But how can you argue about about IP and copyright if the breakthrough of LLMs is potentially based on circumventing or breaking IP and copyright in the first place?
For training the models, and assuming that the content was itself acquired without other acts of infringement? That was ruled legal by the judge in the case I actually (skim) read the judgement of.
At least two companies engaged in acts of infringement to get training data. This is not lawful, and what Anthropic settled out of court for.
As per court ruling, training is fine when they have lawful access.
i.e. open and accessible on the general web is fair game, torrents from the pirate bay is not.
Feel free to argue that copyright law should be changed; this wouldn't be the first time it needed a significant update because new technology made it cheap to do at industrial scale something that was previously so hard that even being able to pull it off made you look legit.
prathje · · focus · HN ↗
We got asked this question two weeks ago when we conducted a workshop on effective and responsible use of AI tools at a local conference. But how can you argue about about IP and copyright if the breakthrough of LLMs is potentially based on circumventing or breaking IP and copyright in the first place?
How do you feel about all of this?
ben_w · · focus · HN ↗
For training the models, and assuming that the content was itself acquired without other acts of infringement? That was ruled legal by the judge in the case I actually (skim) read the judgement of.
At least two companies engaged in acts of infringement to get training data. This is not lawful, and what Anthropic settled out of court for.
prathje · · focus · HN ↗
GJim · · focus · HN ↗
ben_w · · focus · HN ↗
i.e. open and accessible on the general web is fair game, torrents from the pirate bay is not.
Feel free to argue that copyright law should be changed; this wouldn't be the first time it needed a significant update because new technology made it cheap to do at industrial scale something that was previously so hard that even being able to pull it off made you look legit.
desolate_muffin · · focus · HN ↗
ben_w · · focus · HN ↗
I wrote:
> i.e. open and accessible on the general web is fair game, torrents from the pirate bay is not.
Which of these do you think is closer to bypassing a paywall?