GoGuardian
Data Engineer II
United States$130k–$150kmidAdded today
About this role
GoGuardian is seeking a Data Engineer II to design and optimize ETL pipelines, implement MLOps practices, and build scalable data infrastructure supporting analytics and machine learning across the organization. You'll collaborate with data science and BI teams to evolve the data lakehouse platform and ensure data quality and governance.
What you'll do
- Design and optimize ETL pipelines using Databricks, PySpark, and Airflow for analytics and ML workflows
- Develop and maintain labeling and retraining pipelines for machine learning models with quality and reproducibility
- Implement MLOps practices including model versioning, CI/CD, and production monitoring
- Collaborate with data scientists to productionize and scale model training and inference pipelines
- Design and evolve the data lakehouse architecture including schema, partitioning, and performance optimization
- Champion data quality, governance, and documentation to ensure transparency across teams
What they're looking for
- Python and SQL
- PySpark and pandas
- Databricks
- DBT
- Airflow or similar workflow orchestration tools
- AWS data services (S3, Lambda, ECS, CloudWatch)
- Terraform or infrastructure-as-code
- MLOps concepts and practices
Benefits
- Competitive base salary with equity plan
- Comprehensive health insurance and 401(k) matching
- Flexible time off, paid holidays, and paid parental leave
- Work-from-home funds
- Fertility and adoption reimbursement
- Year-end holiday break
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
GoGuardian
- Website
- goguardian.com
Likely interview questions
- Can you describe a complex ETL pipeline you built and how you optimized it for performance and reliability?
- How have you implemented MLOps practices to ensure model reproducibility and monitoring in production?