News publishers such as The Guardian and The New York Times are limiting the Internet Archive's access to their content due to concerns about AI companies scraping their articles. The Internet Archive operates crawlers that capture webpage snapshots, which can be accessed through the Wayback Machine, but this has raised concerns about AI companies using this data to train their models. Some publishers are blocking the Internet Archive's bots or restricting access to their content, while others are working with the Internet Archive to implement changes. This issue highlights the tension between preserving the internet's historical record and protecting intellectual property.