Tech
Wikipedia Thrives in the AI Era, But Researchers Warn of New Challenges from Data Scraping
Wikipedia continues to thrive despite fears that artificial intelligence (AI) would render it obsolete, according to a new study by King’s College London. The research found that the world’s largest online encyclopedia has maintained strong engagement in recent years, even as AI tools like ChatGPT have transformed how people search for and consume information.
Published in the Association for Computing Machinery (ACM) Collective Intelligence journal, the study analyzed data from 12 Wikipedia language editions — six in regions where ChatGPT is available and six where it is not — between January 2021 and January 2024. The findings showed no evidence of declining activity on Wikipedia. In fact, page views and visitor numbers increased across all language editions, although the growth rate was slower in areas where AI chatbots are commonly used.
“We found no signs that ChatGPT reduced the number of Wikipedia editors or edits,” said lead researcher Neal Reeves. However, he noted that some users may have bypassed ChatGPT restrictions using virtual private networks (VPNs), and the study did not account for regional differences in AI adoption.
While the results challenge predictions about the “death of Wikipedia,” the researchers warned that the platform faces serious long-term threats from AI technologies. The report highlights growing issues with “AI scraping” — the large-scale collection of Wikipedia content by AI companies for training their models.
“AI developers are letting their scrapers loose on Wikipedia to train them on high-quality data, pushing traffic to levels where the servers are struggling to keep up,” said Professor Elena Simperl, co-director of the King’s Institute for Artificial Intelligence. She added that AI systems often use Wikipedia content without attribution, diverting web traffic away from the site and undermining its visibility.
Simperl and Reeves urged the creation of a “new social contract” between AI firms and Wikipedia, allowing the encyclopedia to retain control over its data while still supporting responsible AI training.
Coinciding with the study’s release, Wikimedia Deutschland — the German branch of the Wikimedia Foundation — announced the launch of the Wikidata Embedding Project, a new database designed to make Wikipedia content easier and more structured for use in AI systems. The initiative aims to ensure that AI models access verified information curated by Wikipedia editors, rather than relying solely on scraped material.
Researchers say such collaborations could help secure Wikipedia’s relevance in the AI age — balancing open access with the need to protect the integrity and sustainability of one of the internet’s most trusted knowledge sources.
-
Entertainment2 years agoMeta Acquires Tilda Swinton VR Doc ‘Impulse: Playing With Reality’
-
Sports2 years agoChina’s Historic Olympic Victory Sparks National Pride Amid Controversy
-
Business2 years agoSaudi Arabia’s Model for Sustainable Aviation Practices
-
Business2 years agoRecent Developments in Small Business Taxes
-
Home Improvement2 years agoEffective Drain Cleaning: A Key to a Healthy Plumbing System
-
Politics2 years agoWho was Ebrahim Raisi and his status in Iranian Politics?
-
Sports2 years agoKeely Hodgkinson Wins Britain’s First Athletics Gold at Paris Olympics in 800m
-
Business2 years agoCarrectly: Revolutionizing Car Care in Chicago
