Reddit to Update Web Standard to Block Automated Data Scraping From Its Website

Reddit said that researchers and organizations such as the Internet Archive will continue to have access to its content for non-commercial use.

Advertisement
By Reuters | Updated: 26 June 2024 15:51 IST
Highlights
  • AI startups have reportedly been bypassing rules to gather content
  • Reddit said that it would update the Robots Exclusion Protocol
  • The platform will also block unknown bots and crawlers from data scraping

AI firms have been accused of plagiarizing content from publishers

Photo Credit: Reuters

Social media platform Reddit said on Tuesday it will update a Web standard used by the platform to block automated data scraping from its website, following reports that AI startups were bypassing the rule to gather content for their systems.

The move comes at a time when artificial intelligence firms have been accused of plagiarizing content from publishers to create AI-generated summaries without giving credit or asking for permission.

Reddit said that it would update the Robots Exclusion Protocol, or "robots.txt," a widely accepted standard meant to determine which parts of a site are allowed to be crawled.

Advertisement

The company also said it will maintain rate-limiting, a technique used to control the number of requests from one particular entity, and will block unknown bots and crawlers from data scraping - collecting and saving raw information - on its website.

Advertisement

More recently, robots.txt has become a key tool that publishers employ to prevent tech companies from using their content free-of-charge to train AI algorithms and create summaries in response to some search queries.

Last week, a letter to publishers by the content licensing startup TollBit said that several AI firms were circumventing the web standard to scrape publisher sites.

Advertisement

This follows a Wired investigation which found that AI search startup Perplexity likely bypassed efforts to block its Web crawler via robots.txt.

Earlier in June, business media publisher Forbes accused Perplexity of plagiarizing its investigative stories for use in generative AI systems without giving credit.

Advertisement

Reddit said on Tuesday that researchers and organizations such as the Internet Archive will continue to have access to its content for non-commercial use.

© Thomson Reuters 2024


Is the Samsung Galaxy Z Flip 5 the best foldable phone you can buy in India right now? We discuss the company's new clamshell-style foldable handset on the latest episode of Orbital, the Gadgets 360 podcast. Orbital is available on Spotify, Gaana, JioSaavn, Google Podcasts, Apple Podcasts, Amazon Music and wherever you get your podcasts.
Affiliate links may be automatically generated - see our ethics statement for details.
 

For the latest tech news and reviews, follow Gadgets 360 on X, Facebook, WhatsApp, Threads and Google News. For the latest videos on gadgets and tech, subscribe to our YouTube channel. If you want to know everything about top influencers, follow our in-house Who'sThat360 on Instagram and YouTube.

Further reading: Reddit, AI
Advertisement

Related Stories

Popular Mobile Brands
  1. These Poco Phones Will Be Discounted During the Flipkart Big Billion Days
  2. Moto Pad 60 Neo India Launch Date, Key Features, Availability Confirmed
  3. Here's When Your Samsung Galaxy Device Might Get the One UI 8 Update
  4. Xiaomi 15T Series Will Launch With Leica-Tuned Cameras on This Date
  5. Coolie OTT Release Date is Confirmed: All You Need to Know
  6. iPhone 17 Series Launch: Here's a Quick Look at Everything Leaked So Far
  7. Asus VivoBook S16, S14 Price Drops Under Rs. 50,000 With These Offers
  8. iPhone 17 Air, Apple's Slimmest Phone: What to Expect
  9. Apple Could Bring These Major Upgrades to the iPhone 17 Pro Models
  1. Diamond 'Super-Earth' May Not Be Quite as Precious as Once Thought, Study Finds
  2. NASA's James Webb Space Telescope Captures Lobster Nebula’s Towering Spires and Massive Stars
  3. Could a Planet Exist Without a Host Star? Astronomers Say Rogue Worlds May Roam Freely
  4. Exoplanets Explained: How Astronomers Find Worlds Orbiting Stars Beyond the Sun
  5. sPHENIX Detector Clears Test to Study Quark-Gluon Plasma Which Formed After the Big Bang, Claims Study
  6. UY Scuti Reigns as the Universe’s Biggest Known Star, but Its Crown May Be at Risk
  7. Legion Legion Go 2 Will Get ROG Xbox Ally's New Full-Screen Xbox Interface Next Year
  8. Google Nest Cam Outdoor and Indoor Models, Nest Doorbell With Gemini AI Spotted in a Retail Store
  9. Param Sundari OTT Release: When and Where to Watch Janhvi Kapoor-Starrer Online?
  10. Bitcoin’s Largest Whale Dump Since 2022: A Cause for Concern or Just Market Noise?
Gadgets 360 is available in
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2025. All rights reserved.