Skip to main content

Together AI

AI Infrastructure Systems Engineer

San FranciscomidAdded 1 month ago

About this role

Together AI is seeking an AI Infrastructure Engineer in San Francisco to ensure the smooth operation of user-facing services and production systems. The role combines operational expertise and software engineering to enhance system reliability, scalability, and performance.

What you'll do

  • Participate in on-call rotations for production incidents
  • Manage infrastructure using Ansible, Terraform, and Kubernetes
  • Implement monitoring systems for service quality
  • Design operational processes for deployments and upgrades
  • Troubleshoot production issues across services
  • Plan infrastructure growth for Together AI

What they're looking for

  • 5+ years in AI infrastructure or related fields
  • Knowledge of Ansible, Terraform, and Kubernetes
  • Proficient in programming/scripting languages
  • Experience in monitoring and observability
  • Understanding of cloud services
  • Collaborative problem-solving skills

Benefits

  • Competitive compensation
  • Startup equity
  • Health insurance
  • Other competitive benefits
  • Salary range of $190,000 - $270,000
Apply with Autofill

Opens the application — the Jobs AI extension fills it for you. Set up autofill

Opens the official application on the employer’s site. No login required.

Together AI

Together AI builds GPU compute infrastructure and open-source model customization platforms for AI developers and enterprises. The company is hiring for infrastructure operations, ML systems engineering, go-to-market technology, customer success, and GPU research roles.

View all jobs at Together AI