Senior Web Scraping & Data Infrastructure Engineer — Production Systems
Orçamento: $8.0 - $25.0
HOURLY / FULL_TIME
⭐ 5.00 (1)
Pakistan
data-source-integration, data-scraping, etl-pipelines, data-extraction, automation
Qualificações preferidas
- Experiência: Intermédio
We’re looking for an experienced engineer to help us improve and scale an existing web data collection platform.
This is not a project for building a few standalone scraping scripts. We operate a production system that continuously collects and processes data from a large number of websites, including JavaScript-heavy and difficult-to-access sources.
The work will involve both scraping and the infrastructure around it.
You may work on:
* HTTP and browser-based data collection
* Playwright/Puppeteer browser automation
* Proxy, session and cookie management
* Rate limiting and anti-bot challenges
* Distributed workers and queues
* Retries, idempotency and failure recovery
* Scraper monitoring and detecting silent failures
* Data normalization, deduplication and quality checks
* Improving throughput and reducing collection costs
* Designing systems that make adding and maintaining sources easier
Our current environment includes Node.js/TypeScript, Python, Redis, PostgreSQL/MongoDB, Docker, Kubernetes and cloud infrastructure.
We care more about your experience operating scraping systems in production than whether you’ve used every technology in our stack.
We’re looking for someone who can investigate problems independently, explain tradeoffs, and take ownership rather than simply implement tickets.
Abrir na Upwork
AI proposal draft
Generate a short cover letter for this job. Edit before sending.
Sign in to generate an AI proposal draft.
Entrar