OpenAI
Software Engineer, Infrastructure, Consumer Devices
About this role
OpenAI seeks an experienced Cloud Infrastructure Engineer to design and operate scalable platforms powering their consumer products. This hands-on technical leader role combines infrastructure architecture with strategic influence across teams, requiring deep expertise in cloud systems and proven ability to manage large-scale distributed infrastructure.
What you'll do
- Design and build scalable, reliable, secure infrastructure platforms for OpenAI products
- Evolve cloud abstractions to enable rapid product development across multiple teams
- Architect systems supporting significant growth, performance, and operational complexity
- Improve server orchestration, networking, reliability, and infrastructure security
- Participate in on-call rotations, incident response, and production readiness
- Mentor engineers and influence technical direction across the organization
What they're looking for
- 8+ years large-scale infrastructure systems experience
- Deep Kubernetes and container orchestration expertise
- Cloud platform design (AWS, GCP, Azure, or similar)
- Distributed systems reliability and security
- Security engineering background preferred
- Systems thinking and technical leadership
- Cross-team collaboration and mentoring
- Operating in ambiguous, fast-moving environments
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
OpenAI
OpenAI builds AI infrastructure and products, including large-scale data center campuses for AI computing and generative AI applications for enterprise customers. The company is hiring civil engineers, project engineers, electrical design engineers, data center R&D engineers, and AI deployment engineers to expand its infrastructure capabilities and help customers deploy AI solutions.
View all jobs at OpenAILikely interview questions
- Walk us through a time you designed a cloud infrastructure abstraction or platform that enabled multiple teams to move faster. What trade-offs did you make?
- Describe your experience operating Kubernetes at scale. What were the biggest challenges you faced and how did you solve them?