ComingUp ComingUp
W

Website Contact Scraper: Emails, Phone & Social Links

Jul 3, 2026 Marketing & Sales
contact extraction email finder lead generation web scraping

About

Hey r/SideProject, Built this at Techforce Global. The problem: most website contact scrapers use Cheerio (HTML parsing). They are fast and they miss 30–40% of contact data on modern websites because emails, phone numbers, and social links increasingly load via JavaScript after the initial HTML response. I wanted to build a contact scraper that actually handles modern websites correctly. Here is what I built and the technical decisions I made: Playwright instead of Cheerio runs a real Chromium b...

Comments (7)

Mackenzie Conroy Mackenzie Conroy 3 months ago

what's the stack? headless browsers are slow at scale

Aidan Wiza Aidan Wiza 3 months ago

whats the pricing model? per scrape or monthly sub

Damon Kirlin Damon Kirlin 3 months ago

puppeteer or playwright under the hood? speed must take a hit vs cheerio

William Kohler William Kohler 3 months ago

full dom rendering tanks margins at scale.

Harvey Herman Harvey Herman 3 months ago

solving the cheerio gap matters, too many scrapers miss real contact data

Katheryn McCullough Katheryn McCullough 3 months ago

so are you running puppeteer under the hood for the js rendering?

Fernando Kilback Fernando Kilback 2 months ago

what are you parsing with instead of cheerio?