Skip to main content

Indicium AI

AWS Platform Engineer

New York$135k–$190kmidAdded today

About this role

Senior AWS Platform Engineer role focused on designing and scaling cloud infrastructure for AI/ML applications. You'll build production-grade systems using Terraform, Kubernetes (EKS), and GitOps while expanding into MLOps to support model serving and AI workloads at enterprise scale.

What you'll do

  • Design and maintain modular Infrastructure as Code using Terraform across multi-account AWS environments
  • Manage production EKS clusters with Helm charts and dynamic autoscaling (Karpenter/HPA) for CPU/GPU workloads
  • Establish GitOps delivery pipelines (ArgoCD/Flux) for zero-downtime microservices and AI application releases
  • Build cloud infrastructure for ML workflows including model serving, vector databases, and Bedrock/SageMaker integrations
  • Implement IAM policies, network security, secrets management, and observability stacks (Prometheus, Grafana, CloudWatch)
  • Enable developer and data/AI teams by streamlining workflows and standardizing deployment patterns

What they're looking for

  • AWS (multi-account, EKS, Bedrock, SageMaker)
  • Terraform and Infrastructure as Code
  • Kubernetes (EKS) and Helm charts
  • Docker and container orchestration
  • GitOps tools (ArgoCD, Flux)
  • CI/CD automation (GitHub Actions, GitLab CI)
  • Network security and IAM
  • Observability (Prometheus, Grafana, CloudWatch)
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Indicium AI

Indicium AI builds scalable data pipelines and infrastructure that support enterprise AI initiatives, including data integration, warehouse implementation, and ELT processes. The company is hiring Data Engineer Consultants to design and deliver production-grade data solutions while managing cloud infrastructure across cross-functional teams.

View all jobs at Indicium AI

Likely interview questions

  • Walk us through your experience building and managing production EKS clusters at scale—how did you handle autoscaling for mixed CPU/GPU workloads?
  • Describe a complex Terraform infrastructure project you've built; how did you organize it for modularity across multi-account environments?