The Wayback Machine is under threat as major publishers and social media platforms increasingly block its crawlers.
For more than 30 years, the Internet Archive, the nonprofit behind the Wayback Machine, has quietly served as the unofficial chronicler of all things on the web. Since it got its start, it has archived over a trillion pages.
It works by using web crawlers, specifically ia_archiverbot, to scour the web and capture snapshots of sites at different points in time, saving them as a record that internet users can return to.
According to a recent report from WIRED, several legacy and social media companies have begun restricting the Internet Archive’s crawler. That automated retrieval mechanism allows the Wayback Machine to function as the Internet’s memory bank. Organizations named by WIRED include The New York Times and Reddit. Continue reading “Media Companies’ Efforts to Block the Wayback Machine Spur Calls for a Transparent Internet”


