Senior Data Engineer - Scala/Spark
Core
Design, build, and operate large-scale batch data pipelines and infrastructure supporting machine learning and recommendation systems for millions of users globally.
Role type
Senior IC data engineer (Scala/Spark)
Builds
Production-grade data pipelines, ML workflow execution infrastructure, and tooling for data processing
Domain
Technology / Distributed Systems / Machine Learning Infrastructure
Deliverable
production ML models
Required skills
Scala, Apache Spark, distributed data processing, software engineering principles, performance optimization, system reliability, troubleshooting complex systems
Preferred skills
ML infrastructure experience, orchestration platforms, JVM-based distributed systems, cloud/distributed computing environments
Technologies
Scala, Apache Spark
Responsibilities
Design and maintain large-scale Scala and Spark data pipelines; Build new data processing capabilities within an established engineering framework; Own the performance, reliability, and operational health of production data pipelines; Develop tooling and infrastructure supporting ML workflow execution; Optimize data processing systems for throughput and scalability; Collaborate with ML and Research teams to deliver datasets; Identify and resolve system bottlenecks and failures
Seniority
Senior, hands-on IC