{"slug": "stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be", "title": "“Stealth Crawlers” Are Not a Threat to the Open Web. Bills Targeting Them Would Be.", "summary": "The Electronic Frontier Foundation (EFF) warns that proposed legislation targeting so-called \"stealth crawlers\" threatens the open web, user privacy, and valuable research, while the New York state legislature has already passed the NY Stealth Crawler Protection Act, now awaiting Governor Hochul's signature. The EFF argues that anonymous web crawling is essential for investigative journalism, academic research, and cybersecurity, citing examples from The Markup and ProPublica.", "body_md": "There’s a new boogeyman in the battles over AI: so-called “stealth crawlers.” We’ll admit it—the term “stealth crawlers” sound quite nefarious. In reality, they’re anything but.\n\n“Stealth crawlers” are simply automated tools to access and collect public web data—without disclosing the user’s identity. Private crawlers like these facilitate all kinds of [important work](https://www.eff.org/deeplinks/2026/06/free-and-open-web-under-attack-ietf) that benefits the public, including investigative reporting, academic research, cybersecurity protection, and more.\n\nMany publishers want to unmask crawlers anyways—and are pushing for new legislation that would give them new powers to do so. These legislative proposals threaten the open web, user privacy, and valuable research without directly addressing the problems they’re supposedly intending to solve.\n\nAlarmingly, these harmful proposals are gaining traction. The New York state legislature has already passed such a bill, the [NY Stealth Crawler Protection Act](https://www.nysenate.gov/legislation/bills/2025/S9934/amendment/A), which is now on Governor Hochul’s desk. We expect to see similar bills introduced in other states, and potentially in Congress. That’s a big problem for the open web—and the many benefits it provides.\n\nAnonymous crawling enables some of the most publicly beneficial uses of the open web. Researchers, journalists, and other watchdog groups use unidentified automated tools to gather the information necessary to hold powerful institutions accountable and protect the public.\n\nAnonymous crawling fuels important investigative journalism. For example, *The Markup*, a non-profit news site, used [anonymous crawlers](https://themarkup.org/google-the-giant/2020/07/28/how-we-analyzed-google-search-results-web-assay-parsing-tool) to investigate potentially anti-competitive practices by tech companies, such as Amazon’s tendency to [prioritize Amazon brands and Amazon-exclusive products](https://themarkup.org/amazons-advantage/2021/10/14/amazon-puts-its-own-brands-first-above-better-rated-products) over competitors with higher ratings. The crawlers identified themselves as [ordinary Firefox browsers](https://themarkup.org/amazons-advantage/2021/10/14/how-we-analyzed-amazons-treatment-of-its-brands-in-search-results) to web servers, which allowed *The Markup *to understand how Amazon search results pages would appear to ordinary users. Similarly, *ProPublica *used an automated tool designed to simulate an ordinary Amazon customer to reveal that the [site steered shoppers to more expensive products](https://www.propublica.org/article/amazon-says-it-puts-customers-first-but-its-pricing-algorithm-doesnt) over cheaper alternatives.\n\nAnonymous web scraping is also crucial for [cybersecurity professionals](https://www.sans.org/blog/undercover-operations-scraping-the-cybercrime-underground), who use automated tools to monitor the web for information that helps them protect against malicious attackers. [Privacy tools](https://themarkup.org/blacklight), including EFF’s own [Privacy Badger](https://www.eff.org/deeplinks/2023/10/privacy-badger-learns-block-ever-more-trackers), also crawl sites anonymously to identify trackers without compromising user privacy.\n\nHowever, without the ability to scrape anonymously, these tools would likely be blocked. Sites can—and do—block crawlers operated by researchers, journalists, and activists who criticism them. For example, Facebook [shut down](https://www.nytimes.com/2021/08/10/opinion/facebook-misinformation.html?eafs_enabled=false) accounts belonging to researchers who used automated tools to study misinformation on the platform and demanded that they [take down](https://knightcolumbia.org/content/researchers-nyu-knight-institute-condemn-facebooks-effort-to-squelch-independent-research-about-misinformation) published research. Many sites block automated access by anyone who hasn’t paid to crawl public webpages.\n\nNews publishers—and their allies in government—[say that](https://www.newsmediaalliance.org/ny-passes-stealth-crawler-prohibition-act/) unmasking crawlers is necessary to protect news organizations from technological strain caused by AI-related crawling, and fears that AI could reduce news sites’ traffic and ad revenue. These are legitimate concerns.\n\nBut enacting broad, reactionary restrictions on automated access is not the answer. [Legislation targeting anonymous crawling](https://news.bgov.com/bloomberg-government-news/new-york-becomes-first-state-to-ban-bots-that-scrape-news-sites) threatens the open web, user privacy, and valuable research without actually addressing these technological and potential economic harms of scraping.\n\nThe New York state legislature recently passed the [NY Stealth Crawler Protection Act](https://www.nysenate.gov/legislation/bills/2025/S9934/amendment/A), a law that would make it illegal to crawl news websites without revealing who is operating the crawler and all possible future uses of the data collected by the crawler. The law would give websites the power to obtain court orders that unmask anyone using an unidentified crawler—without any evidence that they broke the law.\n\nLaws like the NY bill sweep far beyond AI, and do not meaningfully address the technological or potential harms of AI-related web scraping. These policies would chill beneficial crawling by allowing publishers to veto lawful public access, giving them the power to block not just bad actors, but also security professionals, researchers, dissidents, or anyone who has not paid for a license to view public text. This needlessly undermines the free and open internet.\n\nDigital news publishers—like most websites—face real technological challenges in the AI era. While web crawling has been around for decades, with the proliferation of AI, crawlers now collect far more public web data than they used to. This pushes servers closer to their maximum capacity, and if some bots collect information too aggressively, they may strain web servers to the point that it degrades site performance. The problem is not anonymity—so unmasking crawlers won’t solve it. The real problem is overaggressive crawling, which can be effectively addressed with technical measures that target harmful conduct without impeding anonymous access to information.\n\nThere are other, far less harmful ways to protect publishers from the harms these “stealth crawler” laws claim to target. Addressing the harms of AI-related crawling requires policies that narrowly target the causes of these issues–without undermining free expression and the open web. Policies that target crawlers and scrapers are anything but.", "url": "https://wpnews.pro/news/stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be", "canonical_source": "https://www.eff.org/deeplinks/2026/07/stealth-crawlers-are-not-threat-open-web-bills-targeting-them-would-be", "published_at": "2026-07-20 18:46:50+00:00", "updated_at": "2026-07-20 19:39:28.288429+00:00", "lang": "en", "topics": ["ai-policy", "ai-ethics"], "entities": ["Electronic Frontier Foundation", "NY Stealth Crawler Protection Act", "Governor Hochul", "The Markup", "ProPublica", "Privacy Badger", "Facebook"], "alternates": {"html": "https://wpnews.pro/news/stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be", "markdown": "https://wpnews.pro/news/stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be.md", "text": "https://wpnews.pro/news/stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be.txt", "jsonld": "https://wpnews.pro/news/stealth-crawlers-are-not-a-threat-to-the-open-web-bills-targeting-them-would-be.jsonld"}}