‹ BackHN Continuity

Thread

Spain orders blocks on Archive.today and its mirrors

555 points · 436 comments · latein

  1. peri-cl · · focus · HN ↗
    <a href="https:&#x2F;&#x2F;archive.is&#x2F;2YdN9" rel="nofollow">https:&#x2F;&#x2F;archive.is&#x2F;2YdN9
    1. Lio · · focus · HN ↗
      This is so funny.

      An article about internet censorship of an archive site that requires an archive site to read!

      That’s Brilliant! :D

      1. Imustaskforhelp · · focus · HN ↗
        Let me add a few more layers just for some fun and profit :-D

        Here[0] is an archive.org page which archives an archive.is page which archives the original article about internet censorship of an archive site that might require an archive of an archive site to read (given that people of spain cant now view archive.is in the first place or might have some difficulties doing so)

        [0]: <a href="https:&#x2F;&#x2F;web.archive.org&#x2F;web&#x2F;20260920074416&#x2F;https:&#x2F;&#x2F;serjaimelannister.github.io&#x2F;htmlpipe&#x2F;?https:&#x2F;&#x2F;ppng.io&#x2F;2YdN9_3" rel="nofollow">https:&#x2F;&#x2F;web.archive.org&#x2F;web&#x2F;20260920074416&#x2F;https:&#x2F;&#x2F;serjaimel...

        (If someone is perhaps interested and wants to see me talk more about what this is, then please read the blogpost that I had made for more info: <a href="https:&#x2F;&#x2F;smileplease.mataroa.blog&#x2F;blog&#x2F;htmlpipe-and-how-we-can-use-it-for-archive&#x2F;" rel="nofollow">https:&#x2F;&#x2F;smileplease.mataroa.blog&#x2F;blog&#x2F;htmlpipe-and-how-we-ca...)

        1. 1vuio0pswjnm7 · · focus · HN ↗
          [ERROR] The number of receivers has reached limits.
          1. paul7986 · · focus · HN ↗
            Been using the archive page firefox extension for a year or two. Yet now it looks like a total war against this resource was waged and it&#x27;s no longer a good resource for archiving.

            Anyone know of any good reliable substitutes?

            1. 1vuio0pswjnm7 · · focus · HN ↗
              &quot;Anyone know of any good reliable substitutes?&quot;

              For me, archive.today, archive.is, archive.md, archive.ph, etc. are _not reliable_ for a number of reasons

              But some archive.today users who comment on HN cannot seem to accept that archive.today may not work for everybody else

              NB. Archive.today is not a &quot;substitute for archive.org&quot;. Archive.today does not do www crawls

              As for archive.org, I know of a number of alternatives but each is generally less reliable and&#x2F;or less comprehensive than archive.org

              Comman Crawl, i.e., downloads from data.commoncrawl.org, is reasonably reliable but not as comprehensive as archive.org. CC is not a reasonable substitute for archive.org&#x27;s CDX service. The CC CDX endpoint, index.commoncrawl.org, historically has been easily overwhelmed and unreliable

              As for archive.today alternatives (no crawls, only user-submitted URLs), ghostarchive.org seems well-designed but not used much. No CAPTCHA, HTTPS and Javascript are optional and HAR files are provided. Whether it gets blocked like archive.today sites I do not know

              NB. Archive.today users may be using archive.today not as an archive but as a lazy man&#x27;s solution for &quot;paywalls&quot; (Javascript annoyances)

              Where that&#x27;s the case, comparsions to archive.org or other archives that are derived from crawls are inappropriate

              1. 1vuio0pswjnm7 · · focus · HN ↗
                * Common
Open on Hacker News to reply ↗

Unofficial Hacker News client; not affiliated with Y Combinator.