Clera
Platform Engineer
About this role
Join an AI/ML infrastructure startup as a Platform Engineer to own production reliability, performance, and developer experience across AWS-based systems serving reinforcement learning and agent evaluation workloads. You'll architect and operate containerized infrastructure at scale, leading CI/CD, observability, and incident response for a high-availability platform.
What you'll do
- Own production uptime, latency, provisioning speed, and incident response for core platform services
- Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Docker, and related services
- Design backend systems for scale including capacity planning, autoscaling, queueing, and failure recovery
- Define and improve observability through dashboards, alerts, logs, SLOs, and runbooks
- Build reliable CI/CD pipelines, release automation, and deployment workflows
- Write maintainable code for infrastructure automation and internal tooling
What they're looking for
- AWS (EC2, EKS, S3, ECR, CodeBuild, IAM, networking, secrets management)
- Terraform and infrastructure-as-code
- Kubernetes and container orchestration
- Docker and containerization
- CI/CD pipeline design and automation
- Observability, monitoring, and incident response
- Backend systems architecture and design
- Production systems operations and reliability engineering
Benefits
- Salary: $150,000–$250,000 USD annually
- Equity participation in early-stage, well-funded AI startup
- Work with world-class technical team (competition medalists, founders, researchers)
- High autonomy and direct impact on platform direction
- Visa sponsorship available
- Flexible location options (San Francisco, Singapore, or fully remote)
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Clera
Clera builds an agentic operating system that automates complex workflows and processes through AI agents, with a platform designed to simplify distributed infrastructure management for developers. The company is hiring Founding Engineers, Customer Engineers, and Product Engineers to develop both backend systems and user-facing interfaces across their AI automation products.
View all jobs at CleraLikely interview questions
- Walk us through a production incident you owned end-to-end—what went wrong, how did you detect it, and what did you change to prevent recurrence?
- Describe your approach to capacity planning and autoscaling for a bursty, unpredictable workload. What metrics would you monitor?