Software Engineer, Ingestion Platform
About this role
Reddit's Data Movement team seeks a Software Engineer to build and maintain scalable data infrastructure supporting batch and streaming ML/BI workloads. You'll own the data ingestion platform, work with Spark/Flink/Airflow, and create self-serve systems for hundreds of millions of users generating 100B+ events daily.
What you'll do
- Refine and maintain data infrastructure technologies supporting ML and analytics workflows
- Own the Data Movement Platform for batch and stream data processing
- Build and improve Spark, Flink, and Airflow infrastructure with potential open-source contributions
- Create automated solutions to reduce manual work and enable self-service data access
- Collaborate on on-call rotation, monitoring, and alerting to ensure platform reliability and efficiency
What they're looking for
- Object-oriented programming (Python, Scala, Go, or Java)
- Large-scale system design and implementation
- Airflow orchestration
- Apache Spark
- Apache Flink
- Kubernetes
- CI/CD practices
- Cloud services experience
Benefits
- Comprehensive healthcare and income replacement programs
- 401(k) with employer match
- Flexible vacation and paid volunteer time off
- Generous paid parental leave
- Mental health and coaching benefits
- Gender-affirming care and family planning support
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Reddit builds and operates a large-scale platform serving millions of users, with data infrastructure, advertising systems, and notification products at its core. The company is hiring Software Engineers, Machine Learning Engineers, and Fullstack Engineers to enhance data pipelines, optimize ad delivery, and scale user-facing features.
- Website
- reddit.com
Likely interview questions
- Describe your experience designing and implementing large-scale data systems in production—what were the key challenges and how did you address them?
- Tell us about a time you worked with Spark, Flink, or Airflow at scale. How did you approach optimization and reliability?