Lambda
Data Center Operations Systems Engineer (San Jose)
About this role
Lambda seeks a Data Center Operations Systems Engineer to manage infrastructure deployment, hardware troubleshooting, and inventory operations across AI cloud data centers in San Jose. This on-site shift role involves racking servers, configuring systems, documenting topology, and coordinating with supply chain and support teams to ensure reliable large-scale deployments.
What you'll do
- Rack, cable, label, and configure new server, storage, and network infrastructure
- Troubleshoot hardware and software issues in advanced data center systems
- Document data center layout and network topology using DCIM software
- Coordinate with supply chain and manufacturing teams on deployment timelines and large-scale projects
- Manage parts depot inventory and track equipment through delivery, staging, and deployment cycles
- Collaborate with hardware support and RMA teams to resolve infrastructure issues and manage faulty equipment returns
What they're looking for
- Data center infrastructure systems (power distribution, cooling, environmental monitoring)
- DCIM software and capacity planning
- Structured cabling and cable management
- Server hardware troubleshooting
- Network topology knowledge
- English and Spanish fluency
- Ticketing systems (JIRA, Zendesk)
- Linux administration
Benefits
- Health, dental, and vision coverage for you and dependents
- Generous cash and equity compensation
- 401k plan with 2% company match
- Flexible paid time off
- Wellness and commuter stipends for select roles
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Lambda
Lambda builds AI cloud infrastructure providing GPU compute and networking capabilities for researchers and enterprises. The company is hiring for data center operations, security, and facility engineering roles to support large-scale AI compute deployments.
View all jobs at LambdaLikely interview questions
- Describe your experience with racking, cabling, and configuring server infrastructure in data center environments.
- How have you troubleshot hardware failures in production systems, and what was your approach to documenting and resolving issues?