Software Engineer (Ray Data)
Core
Building and optimizing a Python-native data processing engine (Ray Data) to support scalable AI/ML workloads and multi-modal batch inference.
Role type
Senior distributed systems engineer (data processing)
Builds
Ray Data engine, data loading solutions for production training, scalable AI data pipelines
Domain
Distributed systems, AI/ML infrastructure, data processing
Deliverable
production ML models | infrastructure
Required skills
distributed systems, data processing, database internals, performance optimization, fault tolerance, scaling
Preferred skills
multi-modal data handling, heterogeneous environment scaling, customer-facing technical work
Technologies
Ray, Python
Responsibilities
Improve performance of Ray Data and multi-modal batch inference use cases; Ensure efficient scaling across data pipeline stages in heterogeneous environments; Build data loading solutions for production training workloads; Focus on stability and fault tolerance at high scale; Work with customers and AI-native companies to scale AI workloads
Seniority
Mid-level, hands-on IC