Staff Data Engineer
Core
Designing, building, and maintaining a self-serve data platform (collection, lake management, orchestration, processing, distribution) to support complex analytical and ML workloads.
Role type
Staff Data Engineer (Data Platform)
Builds
High-quality features for a new Data/ML Platform, including batch and streaming data pipelines, storage solutions, and APIs.
Domain
E-commerce / Data Engineering / Cloud Native
Deliverable
production ML models | infrastructure
Required skills
Python, PySpark, SQL, AWS, Databricks (Lakehouse, MLflow, Unity Catalog), Spark applications, Docker, Kubernetes, Terraform, Airflow, data governance, schema management, data validation
Preferred skills
Scala, AWS cost optimization, Mosaic AI, model serving
Technologies
Databricks, AWS, Spark, Python, PySpark, SQL, Docker, Kubernetes, Terraform, Airflow, Parquet, Delta Lake
Responsibilities
Develop and deliver high-quality features for the Data/ML Platform; refactor and translate data products; ensure data quality, schema governance, and monitoring across pipelines; build and maintain Spark applications; manage data collection, lake, orchestration, processing, and distribution.
Seniority
Staff, hands-on IC with high autonomy and impact