An analysis from Originality AI, an AI detection company, found that 23 big news sites block the Internet Archive’s crawler, Wired reports. A USA Today spokesperson told the publication that the move “is not about specifically blocking the Internet Archive” but about attempting to block other scraping bots.
Reddit also limits what the Wayback Machine can archive, telling The Verge last year that it had learned that AI companies were scraping data from the Wayback Machine.













