Staff Software Engineer, Ray Data
Core
Design, build, and optimize the core distributed data processing engine (Ray Data) to power scalable AI workloads, focusing on performance, scalability, and reliability.
Role type
Staff Software Engineer (Distributed Systems & Data Engine)
Builds
Ray Data, a Python-native distributed data processing engine for AI training and inference
Domain
Distributed systems, large-scale data processing, AI infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python engineering, distributed systems internals, system-level architecture, performance optimization, fault tolerance, resource management, batch vs. streaming workloads
Preferred skills
Experience with multimodal data processing, large-scale batch inference, heterogeneous environments
Technologies
Python, Ray
Responsibilities
Design and optimize distributed execution across data pipeline stages; Build data loading and processing solutions for production workloads; Solve problems in scheduling, data partitioning, and consistency; Make system-level architectural decisions on tradeoffs; Work with customers to solve scaling challenges for AI workloads
Seniority
Staff, hands-on IC with strategic impact