Tag: #web-scraping
Writing on GPUs, LLMs, MLOps, Kubernetes — and mindset · 2 posts
Scrapers Now Come From Millions of Home IPs — Residential Proxies and Open-Source Infrastructure
On July 10, 2026, Jonathan Corbet of LWN published an update on the scraper situation. Two facts sit at the center: residential proxy networks have effectively broken IP-based blocking, and the AI-crawler gold rush suppl
2026-07-11 · 6 min read #web-scraping#residential-proxy#ai-crawlers#anubis#infrastructureWeb Scraping & Crawling Tools in 2026 — Scrapy / Playwright / Puppeteer / Crawlee (Apify) / Firecrawl / Jina Reader / Stagehand AI Deep Dive
A single-pass tour of the 2026 web scraping ecosystem. We cover the classics (Scrapy, Playwright, Puppeteer, Selenium), modern frameworks (Crawlee, Apify), proxy clouds (Bright Data, Oxylabs, Smartproxy), API services (S
2026-05-16 · 22 min read #web-scraping#crawling#scrapy#playwright#puppeteer