TensorWave
Storage Platform - Infrastructure Engineer
About this role
TensorWave is seeking a Storage Platform Infrastructure Engineer to maintain and optimize storage systems for Kubernetes and AI/ML workloads. The role involves ensuring the performance and reliability of distributed storage environments while collaborating with various teams to enhance operational efficiency.
What you'll do
- Operate and maintain distributed storage platforms like Ceph and high-performance NAS systems
- Manage storage lifecycle operations including upgrades and migrations
- Monitor and ensure health of storage systems
- Analyze and troubleshoot storage performance metrics
- Support Kubernetes-integrated storage operations
- Enhance automation for storage deployment using tools like Ansible and Terraform
What they're looking for
- 4-7+ years in infrastructure or storage operations
- Experience with distributed storage systems
- Proficient in Ceph and Linux systems
- Knowledge of performance metrics like IOPS and latency
- Troubleshooting skills for storage and network issues
- Experience with Kubernetes and modern storage platforms
- Understanding of data replication
- Familiarity with NVMe storage architectures
Benefits
- [unknown]
- [unknown]
- [unknown]
- [unknown]
- [unknown]
- [unknown]
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
TensorWave
TensorWave builds infrastructure and platforms for large-scale GPU clusters and AI workloads, spanning cloud, data center construction, and virtualization environments. The company is hiring Senior Software Engineers, DevOps Engineers, cloud infrastructure specialists, and operations engineers to develop automation tools, manage complex infrastructure systems, and support high-performance computing across multiple environments.
View all jobs at TensorWave