As part of its mission to preserve the web, the Internet Archive operates crawlers that capture webpage snapshots. As news publishers try to safeguard their contents from AI companies, the Internet Archive is also getting caught in the crosshairs. When asked about The Guardian’s decision, Internet Archive founder Brewster Kahle said that “if publishers limit libraries, like the Internet Archive, then the public will have less access to the historical record.” It’s a prospect, he implied, that could undercut the organization’s work countering “information disorder.”The Guardian isn’t alone in reevaluating its relationship to the Internet Archive. About 93% (226 sites) of publishers in our dataset disallow two out of the four Internet Archive bots we identified. Since there is no federal mandate that requires internet content to be preserved, the Internet Archive is the most robust archiving initiative in the United States.