> Here’s what’s going on. The Internet Archive’s Wayback Machine has been hit by waves of high-volume automated traffic, and we’ve put protections in place to keep the service running.
I'm pretty certain this is scrapers that are trying to workaround blocks on accessing original sites by hitting the Wayback Machine copy instead. Appalling behavior.
In addition to the load it puts on this vital non-profit piece of Internet infrastructure, we've also already seen some sites opt out of the Wayback Machine to prevent their content from being scraped via this alternative route.
This is absolutely something that's happening. There are even paid scraper API's that offer "Wayback Machine fallback" as a feature.
<<a href="https://en.wikipedia.org/wiki/Wikipedia:Archive.today_guidance#Why_are_we_doing_this?" rel="nofollow">https://en.wikipedia.org/wiki/Wikipedia:Archive.today_guidan...> for those without a search engine.
I literally had never heard of this before. I don't check HN every single day.
It's extremely reasonable to ask for a link, very easy to include one when making a claim, and attacking someone for asking for evidence is extremely anti-intellectual independent of the level of effort required.
(n.b. that doesn't excuse the hostile way that they asked for proof - "Let's all just believe this baseless assertion shall we")
They were pointing out the lack of evidence on my part. I agree it was rude but I don't think it's helpful to start calling it misinformation with no evidence. They had a valid point that not everybody Just Knows already, hence why I did reply with a link. I don't think it's constructive to jab much more than I did in that reply.
In some parts of the internet you can't mention a pirate site (or left wing stuff, anything sexual, or Palestine) without being banned. HN isn't one of them, but people have learned to be overly cautious.
simonw · · focus · HN ↗
I'm pretty certain this is scrapers that are trying to workaround blocks on accessing original sites by hitting the Wayback Machine copy instead. Appalling behavior.
In addition to the load it puts on this vital non-profit piece of Internet infrastructure, we've also already seen some sites opt out of the Wayback Machine to prevent their content from being scraped via this alternative route.
packetslave · · focus · HN ↗
bsimpson · · focus · HN ↗
koolala · · focus · HN ↗
sam_lowry_ · · focus · HN ↗
Why being shy in the era of stealing AI?
LoganDark · · focus · HN ↗
schnebbau · · focus · HN ↗
[dead]
LoganDark · · focus · HN ↗
DonHopkins · · focus · HN ↗
[dead]
x______________ · · focus · HN ↗
Wikipedia deprecates Archive.today, starts removing archive links (arstechnica.com) 616 points by nobody9999 6 months ago | hide | past | favorite | 368 comments
0 <a href="https://news.ycombinator.com/item?id=47092006">https://news.ycombinator.com/item?id=47092006
jimmydorry · · focus · HN ↗
1. <a href="https://news.ycombinator.com/item?id=46843805">https://news.ycombinator.com/item?id=46843805
2. <a href="https://news.ycombinator.com/item?id=47092006">https://news.ycombinator.com/item?id=47092006
3. <a href="https://news.ycombinator.com/item?id=47474255">https://news.ycombinator.com/item?id=47474255
gpvos · · focus · HN ↗
DonHopkins · · focus · HN ↗
throw10920 · · focus · HN ↗
It's extremely reasonable to ask for a link, very easy to include one when making a claim, and attacking someone for asking for evidence is extremely anti-intellectual independent of the level of effort required.
(n.b. that doesn't excuse the hostile way that they asked for proof - "Let's all just believe this baseless assertion shall we")
DonHopkins · · focus · HN ↗
[dead]
LoganDark · · focus · HN ↗
throw10920 · · focus · HN ↗
mitxela · · focus · HN ↗