You may have noticed that more and more websites have excluded themselves from the Internet Archive’s invaluable Wayback Machine tool. I certainly have. This Wired article from earlier this year goes into some of the reasons why.

Spoiler: publishers are worried that AI companies are using it to train LLMs on their copyrighted work. I mean, fair enough, but it is a big loss if we gradually lose the web’s historical record in the process.