‹ BackHN Continuity

Thread

An update on Wayback Machine access

685 points · 362 comments · ChrisArchitect

  1. swingandamiss · · focus · HN ↗

    [dead]

    1. righthand · · focus · HN ↗
      No one is upset that the AI companies are scraping the web, they’re upset how poorly implemented the scrapers, but the scraping itself is fine. Lots of people and businesses scrape the web.
      1. akerl_ · · focus · HN ↗
        There are people commenting parallel to you saying they are upset about AI companies scraping the web.
        1. righthand · · focus · HN ↗
          Yeah I dont think they know why they think that.
          1. Joel_Mckay · · focus · HN ↗
            Copyright, patent, and trademark law. The LLM firms misappropriated $9Tn of FOSS community work, ignored the license terms, and resell isomorphic plagiarism tokens to people getting farmed for more data.

            If you still don't understand, than you don't understand how LLM are made.

            Just because something is publicly accessible doesn't mean it is Public Domain. Having hosting people pay for the bots stolen bandwidth (or DDoS), is also theft of service under the law. =3

        2. mitxela · · focus · HN ↗
          Because of the request load though. The ethical thing is separate.
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.