Clinically AI
DevOps Engineer
About this role
Clinically AI seeks a DevOps Engineer to design, build, and operate cloud infrastructure powering its healthcare AI platform. You'll manage GCP, Kubernetes, Terraform, and CI/CD systems while ensuring security, reliability, and scalability for clinical applications and AI workloads.
What you'll do
- Design and maintain GCP infrastructure with GKE, container orchestration, and multi-environment deployments
- Build and evolve Infrastructure-as-Code using Terraform for reproducible, versioned infrastructure
- Develop Helm charts for consistent service deployments, upgrades, and environment management
- Optimize CI/CD pipelines using GitHub Actions for reliable staging, beta, and production deployments
- Implement observability across logging, metrics, tracing, and alerting to monitor system health
- Participate in on-call rotations, incident response, and post-incident reviews
What they're looking for
- Google Cloud Platform (GCP) and GKE
- Terraform and Infrastructure-as-Code
- Kubernetes and container orchestration
- GitHub Actions and CI/CD pipeline design
- Observability, monitoring, and alerting systems
- Infrastructure security and IAM best practices
- Helm and Kubernetes ecosystem tools
- Incident response and reliability engineering
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Clinically AI
Clinically AI builds AI-powered healthcare workflows designed for clinical environments. The company is hiring Quality & Reliability Engineers who specialize in validating AI systems, ensuring operational stability, and managing the unique testing challenges of non-deterministic AI behavior in real-world healthcare settings.
View all jobs at Clinically AILikely interview questions
- Tell us about your experience designing and operating production Kubernetes clusters—what challenges have you faced and how did you solve them?
- Describe your approach to implementing Infrastructure-as-Code with Terraform—how do you handle state management and team collaboration?