New research shows how many important links on the web get lost to time

A quarter of the deep links in The New York Times’ articles are now rotten, leading to completely inaccessible pages, according to a team of researchers from Harvard Law School, who worked with the Times’ digital team. They found that this problem affected over half of the articles containing links in the NYT’s catalog going back to 1996, illustrating the problem of link rot and how difficult it is for context to survive on the web.
The study looked at over 550,000 articles, which contained over 2.2 million links to external websites. It found that 72 percent of those links were “deep,” or pointing to a specific page rather than a general website. Predictably, it found that, as time went on, links were more likely to be dead: 6 percent of links in 2018 articles were inaccessible, while a whopping 72 percent of links from 1998 were dead. For a recent, widespread example of link rot in practice, just look at what happened when Twitter banned Donald Trump: all of the articles that were embedded in his tweets were littered with gray boxes.

The team chose The New York Times in part because the paper is known for its archiving practices, but it’s not suggesting the Times is all that unusual in its link rot problems. Rather, it’s using the paper of record as an example of a phenomenon that happens all across the internet. As time goes by, the websites that once provided valuable insight, important context, or proof of contentious claims through links will be bought and sold, or simply just stop existing, leaving the link to lead to an empty page — or worse.
BuzzFeed News reported in 2019 on the underground industry that exists where customers can pay marketers to find dead links in big outlets like the Times or the BBC and buy the domain for themselves. Then, they can do whatever they want with the link, like using it to advertise products or to host a message making fun of the article’s subject matter.
Link rot doesn’t just affect journalism, either. Imagine if Rick Astley’s “Never Gonna Give You Up” video was deleted and reuploaded. There would be countless Reddit threads and tweet replies that would no longer make sense to future readers. Or imagine if you’re trying to display your NFT, and you discover that the source link now points to nowhere. What a nightmare!
There has been some work done in trying to preserve links. Wikipedia, for example, asks that contributors writing citations provide a link to a page’s archive on sites like the Wayback Machine if they think an article is likely to change. There’s also the Perma.cc project, which attempts to fix the issue of link rot in legal citations and academic journals by providing an archived version of the page, along with a link to the original source.
It’s unlikely, though, that the smattering of similar projects out there would be able to solve the issue for the entire internet, including social networks, or even just for journalists. Until we find a solution, articles will continue to lose more and more context as time goes on. As a perfect example: our article on link rot from 2012 has a source link to The Chesapeake Digital Preservation Group, which now leads to a 404 page.
A quarter of the deep links in The New York Times’ articles are now rotten, leading to completely inaccessible pages, according to a team of researchers from Harvard Law School, who worked with the Times’ digital team. They found that this problem affected over half of the articles containing links…
Recent Posts
- The iOS 18.4 beta brings Matter robot vacuum support
- Philips Monitors is now offering a whopping 5-year warranty on some of its displays, including a gorgeous KVM-enabled business monitor
- The secretive X-37B space plane snapped this picture of Earth from orbit
- Beyond 100TB, here’s how Western Digital is betting on heat dot magnetic recording to reach the storage skies
- The end of an era? TSMC, Broadcom could tear apart Intel’s legendary business after 57 years by separating its foundry and chip design
Archives
- February 2025
- January 2025
- December 2024
- November 2024
- October 2024
- September 2024
- August 2024
- July 2024
- June 2024
- May 2024
- April 2024
- March 2024
- February 2024
- January 2024
- December 2023
- November 2023
- October 2023
- September 2023
- August 2023
- July 2023
- June 2023
- May 2023
- April 2023
- March 2023
- February 2023
- January 2023
- December 2022
- November 2022
- October 2022
- September 2022
- August 2022
- July 2022
- June 2022
- May 2022
- April 2022
- March 2022
- February 2022
- January 2022
- December 2021
- November 2021
- October 2021
- September 2021
- August 2021
- July 2021
- June 2021
- May 2021
- April 2021
- March 2021
- February 2021
- January 2021
- December 2020
- November 2020
- October 2020
- September 2020
- August 2020
- July 2020
- June 2020
- May 2020
- April 2020
- March 2020
- February 2020
- January 2020
- December 2019
- November 2019
- September 2018
- October 2017
- December 2011
- August 2010