extractfeed
A rolling, agent-readable changefeed for web scraping and data extraction.

Legal & Policy

Markdown twin

SerpApi countersues Reddit over API access restrictions

SerpApi has filed counterclaims against Reddit in response to Reddit's lawsuit, alleging broken promises of an open internet and unfair API pricing.

Zyte analysis finds 75% of top sites use robots.txt, but few name specific crawlers

Zyte published a study showing that three-quarters of the world's top websites publish a robots.txt file, yet most do not name individual crawlers.

Zyte argues EU AI scraping guidelines rely on outdated robots.txt standard

Zyte publishes a blog post arguing that Europe's new generative AI scraping guidelines, which lean on the robots.txt protocol, will harm users and entrench monopolies.

Zyte blog post examines how AI and web scraping turn scattered personal data into security risks

Domagoj Marić explores the intersection of AI, web scraping, and OSINT to show how fragmented personal data is assembled into profiles, scams, and security threats at Extract Summit.

Zyte research argues web scraping faces pricing barriers, not outright blocking

Zyte's State of Web Access report, discussed in an interview with the researcher, finds that new economic barriers are making web scraping more difficult rather than technical blocks.