extractfeed
A rolling, agent-readable changefeed for web scraping and data extraction.

Anti-bot & Blocking

Markdown twin

Browserless publishes guide on bypassing Datadome anti-bot protection

Browserless released a blog post detailing techniques to circumvent Datadome's anti-bot system.

Chrome's new navigator.cpuPerformance API opens a fresh fingerprinting vector

Zyte reports that Chrome 152 will expose a navigator.cpuPerformance property, giving sites a new way to fingerprint browsers.

Crawlbase details the infrastructure behind 8,000 CAPTCHAs per second

Crawlbase published a blog post explaining the throughput math, Go-based control plane, and scaling challenges required to solve 8,000 CAPTCHAs per second.

Crawlbase introduces Web Bot Auth as an identity layer for bots, coinciding with pay-per-crawl pricing

Crawlbase ships Web Bot Auth, a specification that treats bot identity as a verifiable signature rather than a self-declared claim, on the same day it launches pay-per-crawl pricing.

Scrapfly publishes 2026 diagnostic guide for blocked Scrapy spiders

Scrapfly released a walkthrough covering nine mechanisms that can block a Scrapy spider, from IP reputation to CAPTCHA, with evidence and mitigation steps for each.

Scrapfly publishes guide on bypassing AWS WAF Bot Control for web scraping

Scrapfly released a blog post detailing how AWS WAF Bot Control detects scrapers across five layers and how to bypass it using their Scrapfly ASP product.

SerpApi explains HTTP 429 errors and rate-limit best practices

SerpApi published a guide covering the causes of HTTP 429 errors, how to read Retry-After headers, and strategies for retrying requests without escalating blocking.

Zyte Audit Reveals Industry-Specific Bot Access Policies

Zyte published a large-scale audit of web access controls showing how different industries enforce different policies toward bots.

Zyte publishes largest ever audit of web access control mechanisms

Zyte released a comprehensive audit of how websites regulate programmatic visits, revealing the current state of web access barriers.

Zyte report finds fashion websites among the most heavily defended against scraping

A Zyte blog post examines the anti-bot and blocking measures used by fashion e-commerce sites, finding they are some of the most aggressively protected on the web.

Zyte report reveals retailers as second most aggressive sector in blocking AI crawlers

Zyte published a blog post detailing how major retail marketplaces deploy anti-bot technology to block AI crawlers and automated data extraction.