No one is upset that the AI companies are scraping the web, they’re upset how poorly implemented the scrapers, but the scraping itself is fine. Lots of people and businesses scrape the web.
Copyright, patent, and trademark law. The LLM firms misappropriated $9Tn of FOSS community work, ignored the license terms, and resell isomorphic plagiarism tokens to people getting farmed for more data.
If you still don't understand, than you don't understand how LLM are made.
Just because something is publicly accessible doesn't mean it is Public Domain. Having hosting people pay for the bots stolen bandwidth (or DDoS), is also theft of service under the law. =3
swingandamiss · · focus · HN ↗
[dead]
righthand · · focus · HN ↗
akerl_ · · focus · HN ↗
righthand · · focus · HN ↗
Joel_Mckay · · focus · HN ↗
If you still don't understand, than you don't understand how LLM are made.
Just because something is publicly accessible doesn't mean it is Public Domain. Having hosting people pay for the bots stolen bandwidth (or DDoS), is also theft of service under the law. =3
mitxela · · focus · HN ↗