Together AI
AI Infrastructure Systems Engineer
San FranciscomidAdded 1 month ago
About this role
Together AI is seeking an AI Infrastructure Engineer in San Francisco to ensure the smooth operation of user-facing services and production systems. The role combines operational expertise and software engineering to enhance system reliability, scalability, and performance.
What you'll do
- Participate in on-call rotations for production incidents
- Manage infrastructure using Ansible, Terraform, and Kubernetes
- Implement monitoring systems for service quality
- Design operational processes for deployments and upgrades
- Troubleshoot production issues across services
- Plan infrastructure growth for Together AI
What they're looking for
- 5+ years in AI infrastructure or related fields
- Knowledge of Ansible, Terraform, and Kubernetes
- Proficient in programming/scripting languages
- Experience in monitoring and observability
- Understanding of cloud services
- Collaborative problem-solving skills
Benefits
- Competitive compensation
- Startup equity
- Health insurance
- Other competitive benefits
- Salary range of $190,000 - $270,000
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Together AI
Together AI builds GPU compute infrastructure and open-source model customization platforms for AI developers and enterprises. The company is hiring for infrastructure operations, ML systems engineering, go-to-market technology, customer success, and GPU research roles.
- Website
- together.ai