Scraping Engineer
New
C
CleraAI / Data Infrastructure
Fully remote — all time zones welcome, provided you can maintain overlap with US business hours., Overlap with US business hoursFull-TimeMiddle
Salary not disclosed
Apply NowOpens the employer's application page
Job Details
- Experience
- 1+ years
- Required Skills
- Node.jsSQLBashGCPGitRabbitmqTypeScriptgRPCRedis
Requirements
- 1+ years of hands-on, professional experience building or maintaining web scraping solutions.
- Proficiency with TypeScript and Node.js for building and debugging scraping scripts.
- Proficiency with SQL for data querying and validation.
- Experience with Puppeteer or similar browser automation libraries.
- Experience debugging and fixing broken scraping scripts in production environments.
- Familiarity with message queues (e.g. RabbitMQ) and asynchronous job processing systems.
- Experience with Redis or similar in-memory caching systems.
- Experience with Google Cloud Platform or equivalent cloud infrastructure.
- Comfort with bash scripting, git, and gRPC.
- Availability for 8+ hours per day with overlap during US business hours.
Responsibilities
- Maintain and monitor a triage queue, resolving broken scrapers and data quality alerts to keep operations running smoothly.
- Build and deploy new web scraping scripts for websites requiring data extraction.
- Validate scraped data for accuracy and investigate discrepancies.
- Create dashboards to visualize and monitor scraped data in real time.
- Collaborate with AI agents to fill capability gaps, fix issues, and improve existing scripts.
View Full Description & ApplyYou'll be redirected to the employer's site