DEV Community

#webscraping

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Why my Reddit scraper went from 92% to 61% success rate in 30 days (and the one-line fix)

Why my Reddit scraper went from 92% to 61% success rate in 30 days (and the one-line fix)

Comments
4 min read
How to Scrape Websites with Claude Code (2026 Guide)

How to Scrape Websites with Claude Code (2026 Guide)

1
Comments
6 min read
Async Python Web Crawler

Async Python Web Crawler

Comments
1 min read
JD.com's isJdSelfRun Flag Is the Best Gray-Market Detection Signal in Chinese E-Commerce (Python Scraper Inside)

JD.com's isJdSelfRun Flag Is the Best Gray-Market Detection Signal in Chinese E-Commerce (Python Scraper Inside)

Comments
6 min read
I stopped treating headless Chrome like a scraping strategy

I stopped treating headless Chrome like a scraping strategy

2
Comments
5 min read
Firecrawl vs Apify vs DivParser: Picking the Right Web Scraping API in 2026

Firecrawl vs Apify vs DivParser: Picking the Right Web Scraping API in 2026

1
Comments
4 min read
6 Firecrawl Alternatives for Developers in 2026

6 Firecrawl Alternatives for Developers in 2026

2
Comments
9 min read
What the Scrapy Maintainer Thinks About AI-Generated Scrapers

What the Scrapy Maintainer Thinks About AI-Generated Scrapers

Comments
2 min read
Building a Letterboxd Film & Review data pipeline: from raw scrape to first insight

Building a Letterboxd Film & Review data pipeline: from raw scrape to first insight

Comments
3 min read
Comparing approaches to extracting Hacker News Who Is Hiring data

Comparing approaches to extracting Hacker News Who Is Hiring data

Comments
3 min read
Sample dataset analysis: a 100-row snapshot of Bazaraki

Sample dataset analysis: a 100-row snapshot of Bazaraki

Comments
3 min read
What I learned scraping Bulk URL Status Checker: schema, gotchas and the tooling that worked

What I learned scraping Bulk URL Status Checker: schema, gotchas and the tooling that worked

Comments
3 min read
What I learned scraping ClinicalTrials.gov: schema, gotchas and the tooling that worked

What I learned scraping ClinicalTrials.gov: schema, gotchas and the tooling that worked

1
Comments 1
3 min read
Build a RAG Pipeline That Actually Reads the Web

Build a RAG Pipeline That Actually Reads the Web

Comments
9 min read
Building a Resilient Wayfair Price Monitor: Python Playwright vs. Node.js Puppeteer

Building a Resilient Wayfair Price Monitor: Python Playwright vs. Node.js Puppeteer

Comments
4 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.