Senior Data Engineer, Python, Spark
Core
Designing data models and building scalable pipelines to capture business metrics across Roku devices, TVs, web, and mobile clients to understand user features and improve experience.
Role type
Senior IC data engineer (big data platform)
Builds
Petabyte-scale data warehouse, distributed data processing systems (batch and streaming), self-service data models
Domain
Streaming media / TV ecosystem / Big Data
Deliverable
production ML models | product features | dashboards & analysis
Required skills
SQL, Python, object-oriented programming, big data technologies (HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, Presto), data modelling, AWS/GCP
Preferred skills
Looker
Technologies
HDFS, YARN, MapReduce, Hive, Kafka, Spark, Airflow, Presto, AWS, GCP, Looker
Responsibilities
Building highly scalable, fault-tolerant distributed data processing systems; Designing and developing robust data solutions; Developing pipelines ensuring high data quality; Defining and maintaining data mappings and transformations; Debugging low-level systems and optimizing clusters; Taking part in architecture discussions and owning new initiatives; Maintaining and evolving existing platforms
Seniority
Senior, hands-on IC