The Wayback Machine Is Blocking Real People by Mistake

An Update on Wayback Machine Access

The Internet Archive says the Wayback Machine has been hit by waves of high-volume automated traffic, forcing new protections to keep it running. Those defenses sometimes flag real users with a 429 "too many requests" error, and the Archive apologizes while it works to better separate abusive bots from the people who depend on the service daily. If you were blocked in error, email info@archive.org with your OS, browser, and IP address.

We're getting better at telling abusive bots apart from the people who depend on the Wayback Machine every day.
  1. simonw

    > Here’s what’s going on. The Internet Archive’s Wayback Machine has been hit by waves of high-volume automated traffic, and we’ve put protections in place to keep the service running.

    I'm pretty certain this is scrapers that are trying to workaround blocks on accessing original sites by hitting the Wayback Machine copy instead. Appalling behavior.

    In addition to the load it puts on this vital non-profit piece of Internet infrastructure, we've also already seen some sites opt out of the Wayback Machine to prevent their content from being scraped via this alternative route.

  2. basilikum

    Mad props to the people at the Archive. You are the heros we need in a formerly open internet that is surrendering to evil big corps and closing down free access.

    The Internet Archive is in a really bad spot being attacked from multiple sides at once. But — while service has not been consistent — they have maintained open access. I can still access anonymously from Tor without Cloudflare or some other centralized gatekeeper showing me the middle finger.

    If you got some money to spare, consider donating to them. They need it.

  3. robotmay

    Unrelated, but this week I've been on a memory binge with the Wayback Machine, trying to find old content of mine from the early 2000s. Took me a while but I've finally put together a good bit of info about myself at the time that I'd completely forgotten, and it's all thanks to the Internet Archive storing my little gaming review website from when I was 16. I could barely remember any of the other stuff, it's been genuinely surprising figuring out what I'd forgotten. I couldn't even remember most domains I owned aside from one, which I used as the starting point.

    Still can't remember what my Tripod site address was, but that might be lost to time.

    Thank you, Archive.org.

  4. userbinator

    The Internet Archive’s Wayback Machine has been hit by waves of high-volume automated traffic

    Thank you for not immediately blaming it on "AI bots". I suspect there's some entity manufacturing consent for strong identity/age verification/sanctioned-browser-OS "walled garden" Internet, and these random DDoSes are part of that.

    I knew something was up when a few alternative YouTube front-ends I use suddenly put up the 'nubis and complained about the high volumes of traffic they were getting flooded with; of course someone actually going after that data would be aiming their "AI bots" at YouTube directly instead of trying to suck it through a tiny little-known proxy-site, so it really strained the credibility of the argument.

  5. BeetleB

    Wow, but I wonder if there's more to it.

    I've not been able to access web.archive.org from my work computer - I always get the 429 error.

    But I then pull out my phone and can access it just fine. All along I was assuming my company was blocking it. Still weird that it happens every time from my work PC and never from my home one.

  6. timpera

    I really appreciate the Archive team's efforts to make the Wayback Machine more responsive, and have donated a few times to support them.

    Unfortunately, the restrictions have been way too strict for the last few months: from my residential IP, simply moving the mouse on the calendar for a specific URL is enough to get stuck on 429 error messages for a while; and from corporate ISPs (for example, on airport WiFi), you often can't access the WM at all. I hope they'll find a way to relax those.

  7. CqtGLRGcukpy

    > We’re getting better at telling abusive bots apart from the people who depend on the Wayback Machine every day. If you think you were blocked in error, email info@archive.org with your operating system, browser, and IP address, and we’ll look into it.

  8. emaro

    It's shame that the AI arms race causes such collateral damage. Free resources were always exploited, but the stakes ($T) and capabilities around AI allow unprecedented abuse. I wish we could go back... :/

    I really don't see any solution to this; the scrapers probably wouldn't even mind destroying sources like IA too much, which would leave them as the only "authorative" source of knowledge in the end. Best way is likely regulation incl. hefty (!) fines, but politics are too slow and too fragmented to be effective. So... Enjoy it while it lasts, I guess.

  9. delis-thumbs-7e

    I recently remembered a wonderful comic blog from 2010’s that is not online anymore. It was a sonderful Finnish LGTG-thened comic blog that I use to read, then forgot completely until few weeks ago. WM had it stored of course, so I could read through this amazing piece of internet art again.

    I really so through some money their way, they do wonderful work.

  10. roughly

    Bonus points for anyone who’d like to guess how the tragedy of the commons was resolved in the times before the enclosure movement.

More from this day

2026-09-15