Wayve
SWE, Data Ingestion
Sunnyvale$210k–$250kfull-timemidAdded today
About this role
Wayve seeks a Data Ingestion Engineer to build and maintain robust, large-scale pipelines that process hundreds of petabytes of autonomous driving data. You'll debug production issues, optimize Spark jobs, and ensure reliable data flow to downstream annotation, data science, and model training teams.
What you'll do
- Debug and resolve failing or blocked ingestion pipelines in production
- Investigate issues caused by corrupt, malformed, or unexpected data formats
- Design resilient pipelines that isolate bad data segments to prevent cascading failures
- Optimize Spark jobs and batch-processing systems for throughput, efficiency, and cost
- Support multi-step workflow orchestration including dependencies, retries, and queue management
- Reduce operational toil by automating manual interventions and improving failure handling
What they're looking for
- Apache Spark (production experience)
- Python engineering
- Large-scale ETL and data-ingestion pipeline design
- Distributed data-processing systems
- Production debugging and troubleshooting
- Data validation and quality assurance
- Workflow orchestration and dependency management
- Cloud infrastructure and cost optimization
Opens the official application on the employer’s site. No login required.
Wayve
Wayve develops autonomous driving AI technology and systems for vehicles. The company is hiring for validation engineers, systems engineers, ML engineers, integration specialists, and program managers to build and deploy autonomous driving platforms across testing, AI development, hardware integration, and production scaling.
- Website
- wayve.ai
Likely interview questions
- Describe a time you debugged a production data pipeline failure at scale—what was the root cause and how did you fix it?
- How would you design a pipeline to handle corrupt or malformed data without blocking downstream teams?