Prompt
I hate how often my scraper gets blocked, what should I switch to?
Latest observation
If you mean “what should I switch to instead of naive scraping,” the best move is usually not another scraper, but a more reliable data-access path:
-
Official API
If the site has one, switch to it. It’s the most stable and least likely to get blocked. -
Licensed data provider / data feed
For business-critical data, paying for access is often cheaper than fighting blocks and breakage. -
Partnership or bulk export
Some sites will provide CSVs, dumps, or partner access if you ask. -
RSS, sitemap, or public feeds
If your use case is updates, these are often easier and lower-friction than page scraping. -
Browser automation only if you must
This is still scraping, but sometimes a headless browser is needed for JS-heavy pages. It can be more fragile, though.
If you’re asking what to switch to technically to reduce blocks without crossing lines, focus on:
- respecting robots.txt and site terms
- slower request rates
- caching and conditional requests
- strong backoff/retry logic
- using identifiable user agents
- limiting concurrency
- keeping sessions consistent
If you want, I can help you choose the right alternative based on what site/data you’re trying to collect and whether this is for personal, research, or business use.
Brands