Klaviyo GMI
Site Reliability Engineer
About this role
Klaviyo seeks a Site Reliability Engineer to design and build highly available, scalable systems that improve service performance and reliability. You'll develop CI/CD infrastructure, collaborate across teams on system optimization, and participate in on-call rotations while championing SRE best practices across the engineering organization.
What you'll do
- Design and develop systems enabling high availability and scalability with focus on performance and efficiency
- Build CI/CD pipeline features and maintain deployment infrastructure for microservices
- Troubleshoot deployment failures, manage database migrations, and support systems during cloud outages
- Conduct quantitative analysis to identify bottlenecks and resolve scalability issues
- Participate in on-call duties to resolve incidents quickly and prevent recurrences
- Collaborate with product teams and advocate for preventative solutions to internal and external stakeholders
What they're looking for
- Python, Java, and Groovy programming
- CI/CD pipeline design and optimization
- Docker and Kubernetes containerization
- AWS services (EC2, Step Functions, Lambda)
- Monitoring tools (New Relic, SumoLogic, EmberJS)
- Database migration management and data security
- Networking concepts and troubleshooting
- Configuration as code and defensive programming
Benefits
- Comprehensive health, welfare, and wellbeing benefits
- Annual cash bonus plan and equity participation
- Telecommuting permitted 2 days per week
- Flexible work arrangement in Boston, MA location
- Professional development and collaborative engineering culture
- Support for responsible AI and workplace accommodations
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Klaviyo GMI
Klaviyo GMI builds scalable data infrastructure and analytics platforms that process billions of weekly events for thousands of customers. The company is hiring Software Engineers, Senior Software Engineers, and Site Reliability Engineers to develop robust backend systems, optimize high-performance databases, build distributed data pipelines, and ensure operational excellence across its platform.
- Website
- klaviyo.com
Likely interview questions
- Describe your experience designing and maintaining CI/CD pipelines for microservices at scale
- How have you identified and eliminated bottlenecks to improve system throughput in a production environment?