Blog

Notes from the scraping trenches.

December 17, 2023

Why we use private proxies for scraping

Public proxy lists look cheap until half of them are already burned. Here's when private proxies are worth paying for — and when they're not.

October 12, 2023

Dealing with CAPTCHAs when you scrape

The best CAPTCHA strategy is not solving them — it's avoiding the behavior that triggers them. When you do hit one, here's the order we try things.

July 3, 2023

When the site ties your session to an IP

Some sites don't just issue a cookie — they expect that cookie to keep coming from the same IP. Rotate carelessly and you'll look logged out, banned, or both.

April 24, 2023

Parsing scraped HTML without making a mess

Fetching the page is the easy half. Turning messy HTML into rows you can trust is where scrapers usually get fragile — and where a few habits save you later.

Need data scraped? Tell us the source — we’ll reply with a plan.

Get in touch