I believe they simply need to integrate a captcha provider and write rules for detecting these captchas on various websites.
Many scraping products and the internal scraping systems of companies (especially retailers) implement such mechanisms.
Many scraping products and the internal scraping systems of companies (especially retailers) implement such mechanisms.
It's honestly a frustrating problem as we're effectively in an arms race with the likes of Cloudflare, Google et al, and we're in the same boat with less reputable services like scrapers and bots. The recent developments around Web Environment Integrity are concerning to us for the same reasons. Feel free to reach out with any questions or suggestions you might have!
A captcha might serve as a strong negative signal.
Disclaimer: I work at CF