OpenAI
Software Engineer, Core Network Engineering
About this role
Join OpenAI's Core Network Engineering team to design and operate the high-performance networking infrastructure powering large-scale AI training and inference systems. You'll work across datacenter fabrics, host networking, and global WAN systems where microseconds of latency directly impact model efficiency.
What you'll do
- Design and operate networking systems supporting large-scale AI training and inference infrastructure
- Optimize performance and reliability across host networking, datacenter fabrics, and WAN technologies
- Develop automation for provisioning, configuration management, and infrastructure lifecycle operations
- Build observability and diagnostic tooling for network health, performance analysis, and remediation
- Optimize high-performance networking technologies including RDMA, RoCE, InfiniBand, and GPU interconnects
- Partner with compute, storage, and hardware teams on architecture, capacity planning, and reliability
What they're looking for
- Linux networking and kernel systems
- High-performance networking technologies (InfiniBand, RoCE, DPDK, Ethernet fabrics)
- Datacenter or WAN networking architecture
- C++, Python, or Go
- RDMA and NIC optimization
- Distributed systems debugging and performance analysis
- Infrastructure automation and configuration management
- Systems fundamentals (networking, OS, distributed systems)
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
OpenAI
OpenAI builds AI infrastructure and products, including large-scale data center campuses for AI computing and generative AI applications for enterprise customers. The company is hiring civil engineers, project engineers, electrical design engineers, data center R&D engineers, and AI deployment engineers to expand its infrastructure capabilities and help customers deploy AI solutions.
View all jobs at OpenAILikely interview questions
- Tell us about your experience building or operating large-scale networking infrastructure. What was the most significant performance bottleneck you diagnosed and resolved?
- Describe your hands-on experience with high-performance networking technologies like RDMA, RoCE, InfiniBand, or DPDK. How have you optimized network performance in production systems?