Clera
Scraping Engineer
About this role
Own the reliability of a web scraping infrastructure processing millions of pages daily at an AI startup. You'll monitor and fix broken scrapers, build new extraction scripts, validate data quality, and collaborate with AI agents to keep a high-throughput pipeline running smoothly.
What you'll do
- Monitor and triage broken scrapers and data quality alerts to maintain pipeline operations
- Build and deploy new web scraping scripts for various target websites
- Validate scraped data for accuracy and investigate discrepancies
- Create dashboards to visualize and monitor scraped data in real time
- Collaborate with AI agents to fix issues and improve existing scripts
- Debug and resolve production scraper failures
What they're looking for
- TypeScript and Node.js
- Web scraping and browser automation (Puppeteer)
- SQL for data querying and validation
- Message queues (RabbitMQ or similar)
- Redis or in-memory caching systems
- Google Cloud Platform and BigQuery
- Bash scripting and git
- Data visualization and dashboard tools
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Clera
Clera builds an agentic operating system that automates complex workflows and processes through AI agents, with a platform designed to simplify distributed infrastructure management for developers. The company is hiring Founding Engineers, Customer Engineers, and Product Engineers to develop both backend systems and user-facing interfaces across their AI automation products.
View all jobs at CleraLikely interview questions
- Describe a time when you debugged a broken scraper in production—what was the issue and how did you resolve it?
- How would you validate scraped data to ensure accuracy when the underlying website structure changes?