Scraping Engineer

New
C
CleraAI / Data Infrastructure
Fully remote — all time zones welcome, provided you can maintain overlap with US business hours., Overlap with US business hoursFull-TimeMiddle
Salary not disclosed
Apply NowOpens the employer's application page

Job Details

Experience
1+ years
Required Skills
Node.jsSQLBashGCPGitRabbitmqTypeScriptgRPCRedis

Requirements

  • 1+ years of hands-on, professional experience building or maintaining web scraping solutions.
  • Proficiency with TypeScript and Node.js for building and debugging scraping scripts.
  • Proficiency with SQL for data querying and validation.
  • Experience with Puppeteer or similar browser automation libraries.
  • Experience debugging and fixing broken scraping scripts in production environments.
  • Familiarity with message queues (e.g. RabbitMQ) and asynchronous job processing systems.
  • Experience with Redis or similar in-memory caching systems.
  • Experience with Google Cloud Platform or equivalent cloud infrastructure.
  • Comfort with bash scripting, git, and gRPC.
  • Availability for 8+ hours per day with overlap during US business hours.

Responsibilities

  • Maintain and monitor a triage queue, resolving broken scrapers and data quality alerts to keep operations running smoothly.
  • Build and deploy new web scraping scripts for websites requiring data extraction.
  • Validate scraped data for accuracy and investigate discrepancies.
  • Create dashboards to visualize and monitor scraped data in real time.
  • Collaborate with AI agents to fill capability gaps, fix issues, and improve existing scripts.
View Full Description & ApplyYou'll be redirected to the employer's site
View details
Apply Now