Anyscale
Software Engineer (Ray Data)
About this role
Anyscale is seeking a Software Engineer to develop and optimize Ray Data, a Python-native distributed data processing engine for AI/ML workloads. You'll focus on performance improvements, scaling efficiency, fault tolerance, and building solutions that help customers run complex AI applications at scale.
What you'll do
- Improve Ray Data performance for multi-modal batch inference use cases
- Optimize data pipeline scaling across heterogeneous environments
- Build data loading solutions for production training workloads
- Enhance stability and fault tolerance at high scale
- Collaborate with customers to scale their AI workloads
- Maintain and develop Ray Data distributed processing engine
What they're looking for
- Distributed systems design and implementation
- Python programming
- Data processing and database internals
- Performance optimization
- Fault tolerance and reliability engineering
- Large-scale system architecture
- AI/ML infrastructure understanding
- Scalability design patterns
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Anyscale
Anyscale builds Ray, an open-source distributed computing framework and enterprise platform for scaling AI workloads across Kubernetes and cloud providers. The company is hiring forward-deployed engineers to work embedded with customers, software engineers to develop Ray Core, LLM inference specialists, and customer support engineers who combine technical expertise with post-sale success.
- Website
- anyscale.com
Likely interview questions
- Describe your experience building or optimizing distributed data processing systems at scale
- How have you approached performance optimization in a heterogeneous computing environment?