Link Rot and Digital Decay on Government, News and Other Webpages – Pew Research Center
Pew Research Center conducted the analysis to examine how often online content that once existed becomes inaccessible. One part of the study looks at a representative sample of webpages that existed over the past decade to see how many are still accessible today. For this analysis, we collected a sample of pages from the Common Crawl web repository for each year from 2013 to 2023. We then tried to access those pages to see how many still exist.
A second part of the study looks at the links on existing webpages to see how many of those links are still functional. We did this by collecting a large sample of pages from government websites, news websites and the online encyclopedia Wikipedia.
We identified relevant news domains using data from the audience metrics company comScore and relevant government domains (at multiple levels of government) using data from get.gov, the official administrator for the .gov domain. We collected the news and government pages via Common Crawl and the Wikipedia pages from an archive maintained by the Wikimedia Foundation. For each collection, we identified the links on those pages and followed them to their destination to see what share of those links point to sites that are no longer accessible.
A third part of the study looks at how often individual posts on social media sites are deleted or otherwise removed from public view. We did this by collecting a large sample of public tweets on the social media platform X (then known as Twitter) in real time using the Twitter Streaming API. We then tracked the status of those tweets for a period of three months using the Twitter Search API to monitor how many were still publicly available. Refer to the report methodology for more details.
The internet is an unimaginably vast repository of modern life, with hundreds of billions of indexed webpages. But even as users across the world rely on the web to access books, images, news articles and other resources, this content sometimes disappears from view.
A new Pew Research Center analysis shows just how fleeting online content actually is:
This digital decay occurs in many different online spaces. We examined the links that appear on government and news websites, as well as in the References section of Wikipedia pages as of spring 2023. This analysis found that:
To see how digital decay plays out on social media, we also collected a real-time sample of tweets during spring 2023 on the social media platform X (then known as Twitter) and followed them for three months. We found that:
There are many ways of defining whether something on the internet that used to exist is now inaccessible to people trying to reach it today. For instance, inaccessible could mean that:
For this report, we focused on the first of these: pages that no longer exist. The other definitions of accessibility are beyond the scope of this research.
Our approach is a straightforward way of measuring whether something online is accessible or not. But even so, there is some ambiguity.
First, there are dozens of status codes indicating a problem that a user might encounter when they try to access a page. Not all of them definitively indicate whether the page is permanently defunct or just temporarily unavailable. Second, for security reasons, many sites actively try to prevent the sort of automated data collection that we used to test our full list of links.
For these reasons, we used the most conservative estimate possible for deciding whether a site was actually accessible or not. We counted pages as inaccessible only if they returned one of nine error codes that definitively indicate that the page and/or its host server no longer exist or have become nonfunctional regardless of how they are being accessed, and by whom. The full list of error codes that we included in our definition are in the methodology.
Here are some of the findings from our analysis of digital decay in various online spaces.
To conduct this part of our analysis, we collected a random sample of just under 1 million webpages from the archives of Common Crawl, an internet archive service that periodically collects snapshots of the internet as it exists at different points in time. We sampled pages collected by Common Crawl each year from 2013 through 2023 (approximately 90,000 pages per year) and checked to see if those pages still exist today.
We found that 25% of all the pages we collected from 2013 through 2023 were no longer accessible as of October 2023. This figure is the sum of two different types of broken pages: 16% of pages are individually inaccessible but come from an otherwise functional root-level domain; the other 9% are inaccessible because their entire root domain is no longer functional.
Not surprisingly, the older snapshots in our collection had the largest share of inaccessible links. Of the pages collected from the 2013 snapshot, 38% were no longer accessible in 2023. But even for pages collected in the 2021 snapshot, about one-in-five were no longer accessible just two years later.
We sampled around 500,000 pages from government websites using the Common Crawl March/April 2023 snapshot of the internet, including a mix of different levels of government (federal, state, local and others). We found every link on each page and followed a random selection of those links to their destination to see if the pages they refer to still exist.
Across the government websites we sampled, there were 42 million links. The vast majority of those links (86%) were internal, meaning they link to a different page on the same website. An explainer resource on the IRS website that links to other documents or forms on the IRS site would be an example of an internal link.
Around three-quarters of government webpages we sampled contained at least one on-page link. The typical (median) page contains 50 links, but many pages contain far more. A page in the 90th percentile contains 190 links, and a page in the 99th percentile (that is, the top 1% of pages by number of links) has 740 links.
Other facts about government webpage links:
When we followed these links, we found that 6% point to pages that are no longer accessible. Similar shares of internal and external links are no longer functional.
Overall, 21% of all the government webpages we examined contained at least one broken link. Across every level of government we looked at, there were broken links on at least 14% of pages; city government pages had the highest rates of broken links.
For this analysis, we sampled 500,000 pages from 2,063 websites classified as News/Information by the audience metrics firm comScore. The pages were collected from the Common Crawl March/April 2023 snapshot of the internet.
Across the news sites sampled, this collection contained more than 14 million links pointing to an outside website. Some 94% of these pages contain at least one external-facing link. The median page contains 20 links, and pages in the top 10% by link count have 56 links.
Like government websites, the vast majority of these links go to secure HTTP pages (those with a URL beginning with https://). Around 12% of links on these news sites point to a static file, like a PDF document. And 32% of links on news sites redirected to a different URL than the one they originally pointed to slightly less than the 39% of external links on government sites that redirect.
When we tracked these links to their destination, we found that 5% of all links on news site pages are no longer accessible. And 23% of all the pages we sampled contained at least one broken link.
Broken links are about as prevalent on the most-trafficked news websites as they are on the least-trafficked sites. Some 25% of pages on news websites in the top 20% by site traffic have at least one broken link. That is nearly identical to the 26% of sites in the bottom 20% by site traffic.
For this analysis, we collected a random sample of 50,000 English-language Wikipedia pages and examined the links in their References section. The vast majority of these pages (82%) contain at least one reference link that is, one that directs the reader to a webpage other than Wikipedia itself.
In total, there are just over 1 million reference links across all the pages we collected. The typical page has four reference links.
The analysis indicates that 11% of all references linked on Wikipedia are no longer accessible. On about 2% of source pages containing reference links, every link on the page was broken or otherwise inaccessible, while another 53% of pages contained at least one broken link.
For this analysis, we collected nearly 5 million tweets posted from March 8 to April 27, 2023, on the social media platform X, which at the time was known as Twitter. We did this using Twitters Streaming API, collecting 3,000 public tweets every 30 minutes in real time. This provided us with a representative sample of all tweets posted on the platform during that period. We monitored those tweets until June 15, 2023, and checked each day to see if they were still available on the site or not.
At the end of the observation period, we found that 18% of the tweets from our initial collection window were no longer publicly visible on the site. In a majority of cases, this was because the account that originally posted the tweet was made private, suspended or deleted entirely. For the remaining tweets, the account that posted the tweet was still visible on the site, but the individual tweet had been deleted.
Tweets were especially likely to be deleted or removed over the course of our collection period if they were:
We also found that removed or deleted tweets tended to come from newer accounts with relatively few followers and modest activityon the site. On average, tweets that were no longer visible on the site were posted by accounts around eight months younger than those whose tweets stayed on the site.
And when we analyzed the types of tweets that were no longer available, we found that retweets, quote tweets and original tweets did not differ much from the overall average. But replies were relatively unlikely to be removed just 12% of replies were inaccessible at the end of our monitoring period.
Most tweets that are removed from the site tend to disappear soon after being posted. In addition to looking at how many tweets from our collection were still available at the end of our tracking period, we conducted a survival analysis to see how long these tweets tended to remain available. We found that:
Put another way: Half of tweets that are eventually removed from the platform are unavailable within the first six days of being posted. And 90% of these tweets are unavailable within 46 days.
Tweets dont always disappear forever, though. Some 6% of the tweets we collected disappeared and then became available again at a later point. This could be due to an account going private and then returning to public status, or to the account being suspended and later reinstated. Of those reappeared tweets, the vast majority (90%) were still accessible on Twitter at the end of the monitoring period.
Link:
Link Rot and Digital Decay on Government, News and Other Webpages - Pew Research Center
- Elon Musks Anti-Woke Wikipedia Is Calling Hitler The Fhrer - The Intercept - November 30th, 2025 [November 30th, 2025]
- 50 Times People Found Gems On Wikipedia That Were Too Funny Not To Share (New Pics) - Bored Panda - November 30th, 2025 [November 30th, 2025]
- For Wikipedia founder Jimmy Wales, truth has always been a matter of trust | The Excerpt - USA Today - November 30th, 2025 [November 30th, 2025]
- How does the Wikimedia Foundation use donations to Wikipedia? - Wikimedia Foundation - November 30th, 2025 [November 30th, 2025]
- The Interview: How Wikipedia Is Responding to the Culture Wars - The New York Times - November 30th, 2025 [November 30th, 2025]
- For Wikipedia founder Jimmy Wales, truth is a matter of trust - USA Today - November 30th, 2025 [November 30th, 2025]
- ShellBot Chat: How to Edit History The Wikipedia Way - Royal Dutch Shell Plc .com - November 30th, 2025 [November 30th, 2025]
- Grok, is this true? Can Elon Musk's Grokipedia compete with Wikipedia? - Mezha - November 30th, 2025 [November 30th, 2025]
- The best guide to spotting AI writing comes from Wikipedia - TechCrunch - November 20th, 2025 [November 20th, 2025]
- The difference between Grokipedia and Wikipedia - marketplace.org - November 20th, 2025 [November 20th, 2025]
- 27 Wikipedia Pages So Disturbing They're For Adults Only - BuzzFeed - November 20th, 2025 [November 20th, 2025]
- Wikipedia founder Jimmy Wales blows his top and hits da bricks 45 seconds into an interview, shouting 'It's a stupid question!' as he walks offstage -... - November 20th, 2025 [November 20th, 2025]
- Wikipedia Cracks the Code on Spotting AI Writing - The Tech Buzz - November 20th, 2025 [November 20th, 2025]
- Elon Musk, Wikipedia co-founder Jimmy Wales is not pleased with your Wikipedia rival; says: Pretty skepti - Times of India - November 20th, 2025 [November 20th, 2025]
- As Wikipedia Traffic Drops 8%, Experts Say Its Time to Rethink SEO and GEO - DesignRush - November 20th, 2025 [November 20th, 2025]
- I Really, Really, Really, Really, Really, Really, Really, Really, Really, Really, Really, Really Regret Looking At These Creepy Wikipedia Pages -... - November 20th, 2025 [November 20th, 2025]
- Jimmy Wales walks out of interview over dumbest Wikipedia question: Its not a - Times of India - November 20th, 2025 [November 20th, 2025]
- Wikipedia is facing attacks from the White House and Musk. Its founder isn't worried - NPR - November 7th, 2025 [November 7th, 2025]
- We tried Elon Musks Wikipedia clone. Its as racist as youd expect - The Sydney Morning Herald - November 7th, 2025 [November 7th, 2025]
- We tried Elon Musks Wikipedia clone. Its as racist as youd expect - The Sydney Morning Herald - November 7th, 2025 [November 7th, 2025]
- Ranked: The Most Viewed Wikipedia Pages of 2025 (So Far) - Visual Capitalist - November 7th, 2025 [November 7th, 2025]
- Ranked: The Most Viewed Wikipedia Pages of 2025 (So Far) - Visual Capitalist - November 7th, 2025 [November 7th, 2025]
- I Fell Into The Darkest Parts Of Wikipedia And I Want A Refund - BuzzFeed - November 7th, 2025 [November 7th, 2025]
- I Fell Into The Darkest Parts Of Wikipedia And I Want A Refund - BuzzFeed - November 7th, 2025 [November 7th, 2025]
- How Wikipedia co-founder Jimmy Wales may have agreed with Elon Musk that Wikipedia is 'biased' - The Times of India - November 7th, 2025 [November 7th, 2025]
- We tried Elon Musks Wikipedia clone. Its as racist as youd expect - The Age - November 7th, 2025 [November 7th, 2025]
- INSEAD launches Botipedia, an AI-created encyclopedic knowledge portal that claims to be 6,000 times larger than Wikipedia - EdTech Innovation Hub - November 7th, 2025 [November 7th, 2025]
- I tried Elon Musk's Wikipedia clone and boy is it racist - SFGATE - November 5th, 2025 [November 5th, 2025]
- Elon Musk? AI? Crazy left-wing activists? The main who built Wikipedia explains its biggest threats - BBC Science Focus Magazine - November 5th, 2025 [November 5th, 2025]
- Musk version of Wikipedia takes different tack on climate - E&E News by POLITICO - November 5th, 2025 [November 5th, 2025]
- I tried Grokipedia. It has something to teach Wikipedia about AI. - Business Insider - November 3rd, 2025 [November 3rd, 2025]
- Step aside, Wikipedia; its Grok to the future - Washington Times - November 3rd, 2025 [November 3rd, 2025]
- AI answers are taking a bite of Wikipedia's traffic. Should we be worried for the site? - Business Insider - November 3rd, 2025 [November 3rd, 2025]
- Wikipedia sends 'note' to everyone on the internet as it takes on Elon Musk's Grokipedia - The Times of India - November 3rd, 2025 [November 3rd, 2025]
- What Elon Musks Version of Wikipedia Thinks About Hitler, Putin, and Apartheid - The Atlantic - November 3rd, 2025 [November 3rd, 2025]
- I tried Grokipedia, the AI-powered anti-Wikipedia. Here's why neither is foolproof - ZDNET - November 3rd, 2025 [November 3rd, 2025]
- Why Wikipedia Is Losing Traffic to AI Overviews on Google - CNET - November 3rd, 2025 [November 3rd, 2025]
- Grokipedia vs Wikipedia: How Elon Musk's AI-generated encyclopaedia holds up against the left-leaning cro - The Times of India - November 3rd, 2025 [November 3rd, 2025]
- WIKIPEDIA CO-FOUNDER: WIKIPEDIA WILL BE LEFT IN THE DUST BY GROKIPEDIA" Ex-founder of Wikipedia, Larry Sanger: "The neat thing that theyre... - November 3rd, 2025 [November 3rd, 2025]
- How AI could soon be used by Wikipedia, according to its founder - BBC Science Focus Magazine - November 3rd, 2025 [November 3rd, 2025]
- Grokipedia Is the Antithesis of Everything That Makes Wikipedia Good, Useful, and Human - 404 Media - November 3rd, 2025 [November 3rd, 2025]
- Seth Meyers Drags Trump for Having an Entire Wikipedia Page Dedicated to His Handshake Technique | Video - TheWrap - November 3rd, 2025 [November 3rd, 2025]
- Elon Musk Launches AI-Powered Rival to Wikipedia and Its Already Been Accused of Copying Wiki Pages - People.com - November 3rd, 2025 [November 3rd, 2025]
- Wikipedia says AI answers are starting to take a bite. There are reasons to be worried. - Yahoo News Canada - November 3rd, 2025 [November 3rd, 2025]
- What Wikipedia and Grokipedia are saying about each other - KGOU - November 3rd, 2025 [November 3rd, 2025]
- I pitted Wikipedia against Elon Musks new Grokipedia heres which one gave the better answers - Tom's Guide - November 3rd, 2025 [November 3rd, 2025]
- Explained | What is Grokipedia, Musk's AI alternative to human-edited Wikipedia - Deccan Herald - November 3rd, 2025 [November 3rd, 2025]
- AI still cant beat Wikipedia when it comes to integrity - The Observer - November 3rd, 2025 [November 3rd, 2025]
- Elon Musk's 'Grokipedia' cites Wikipedia as a source, even though it's the exact thing he's trying to replace because he thinks it's 'woke' - Fortune - November 3rd, 2025 [November 3rd, 2025]
- WIKIPEDIA TRIED TO ROAST GROKIPEDIA AND COOKED ITS OWN CREDIBILITY In a new fundraising pop-up, Wikipedia throws shade at Grokipedia, bragging it's... - November 3rd, 2025 [November 3rd, 2025]
- Elon Musk wants to dethrone Wikipedia with Grokipedia - MSN - November 3rd, 2025 [November 3rd, 2025]
- Grokipedia: Far right talking points or much-needed antidote to Wikipedia? - TradingView - November 3rd, 2025 [November 3rd, 2025]
- Hi, Its Me, Wikipedia, and I Am Ready for Your Apology - McSweeneys Internet Tendency - October 28th, 2025 [October 28th, 2025]
- Watch Wikipedia Founder Wales Explores Trust in the Digital Age - Bloomberg.com - October 28th, 2025 [October 28th, 2025]
- He co-founded Wikipedia. Now hes inspiring Elon Musk to build a rival. - Yahoo - October 28th, 2025 [October 28th, 2025]
- 'An astonishing situation': Wikipedia co-founder bashes Trump's latest attacks on trust - rawstory.com - October 28th, 2025 [October 28th, 2025]
- Trust and empathy should be baked into tech from the start, says Wikipedia co-founder - marketplace.org - October 28th, 2025 [October 28th, 2025]
- Elon Musks Grokipedia copying Wikipedia? Here's all you need to know about the AI-powered encyclopedia - The Economic Times - October 28th, 2025 [October 28th, 2025]
- Explained: What is Elon Musks Grokipedia and how it differs from Wikipedia - The Federal - October 28th, 2025 [October 28th, 2025]
- Grokipedia Vs Wikipedia: How Is The Elon Musk's AI-Powered Rival Different From The Encyclopedia? - Mashable India - October 28th, 2025 [October 28th, 2025]
- Elon Musks xAI launches AI-powered Grokipedia database to replace Wikipedia - The Hindu - October 28th, 2025 [October 28th, 2025]
- Grokipedia is online: Elon Musk's AI encyclopedia wants to crush Wikipedia - Cointribune - October 28th, 2025 [October 28th, 2025]
- Elon Musks Grokipedia Takes Aim at Wikipedia Truth Revolution or Biased Echo Chamber? - ts2.tech - October 28th, 2025 [October 28th, 2025]
- Elon Musks Version of Wikipedia Is Live. Heres What the Difference Is - Gizmodo - October 28th, 2025 [October 28th, 2025]
- Even Grokipedia needs Wikipedia to exist: Is Elon Musk's AI-powered encyclopedia less biased as he claims? - theweek.in - October 28th, 2025 [October 28th, 2025]
- Elon Musks Wikipedia Alternative Grokipedia Goes Live: Heres How To Use It - NDTV Profit - October 28th, 2025 [October 28th, 2025]
- Cry Us a River: AI Chatbots May Be Killing Wikipedia - Science and Culture Today - October 28th, 2025 [October 28th, 2025]
- Elon Musk launches rival to challenge Wikipedia; Here's all you need to know about this - DNA India - October 28th, 2025 [October 28th, 2025]
- GROKIPEDIA IS ALREADY MORE ACCURATE THAN WIKIPEDIA AND IT SHOWS Grokipedia just proved why it is rewriting how knowledge works online. Look at how it... - October 28th, 2025 [October 28th, 2025]
- Nothing But The Truth: Will Elon Musk's Grokipedia Deal A Death Blow To 'Woke' Wikipedia? - News18 - October 28th, 2025 [October 28th, 2025]
- Grokipedia launched by Elon Musk to take on Wikipedia: Heres how to use it, new AI features, early controversy, and more - financialexpress.com - October 28th, 2025 [October 28th, 2025]
- Grokipedia Debuts: Elon Musks AI-Powered Alternative to Wikipedia - parameter.io - October 28th, 2025 [October 28th, 2025]
- The Wikipedia Page on "Brain Rot" Is Protected Until 2026 Due to Extensive Vandalism - Futurism - October 26th, 2025 [October 26th, 2025]
- 'I was very nervous at first' - how the founder of Wikipedia learnt to embrace trust - RNZ - October 26th, 2025 [October 26th, 2025]
- A Wikipedia cofounder is fueling the rights campaign against it - The Washington Post - October 24th, 2025 [October 24th, 2025]
- Where does Wikipedia go in the age of AI? - Financial Times - October 24th, 2025 [October 24th, 2025]
- Wikipedia co-founder Larry Sangers long-standing claims of liberal bias and mismanagement at the worlds dominant online encyclopedia are being... - October 24th, 2025 [October 24th, 2025]
- Grokipedia was supposed to rival Wikipedia but Elon Musk pulled the plug (for now) - Tom's Guide - October 24th, 2025 [October 24th, 2025]
- Murdaugh: Death In The Family Owes More Than You Think To One Wikipedia Line - Screen Rant - October 24th, 2025 [October 24th, 2025]
- Wikipedia blames ChatGPT for falling traffic and claims bots are stealing its hard work - New York Post - October 24th, 2025 [October 24th, 2025]