Lead Data Engineer
Core
Design, build, and operate reliable, scalable, and observable data pipelines that power business functions across the enterprise.
Role type
Senior individual-contributor data engineer (Databricks/PySpark)
Builds
Production-grade ETL/ELT pipelines, reusable frameworks for ingestion and transformation, curated datasets for downstream analytics
Domain
Enterprise data engineering, cloud platforms (AWS preferred)
Deliverable
production ML models | product features
Required skills
Databricks, Python (PySpark), SQL, CI/CD, Git-based branching strategies, cloud platforms (AWS), monitoring and alerting, incident response, data quality management, pipeline architecture
Preferred skills
Orchestration frameworks, streaming technologies, Infrastructure-as-Code, observability tooling, semiconductor manufacturing or large-scale industrial data processing background
Technologies
Databricks, PySpark, AWS, Git
Responsibilities
Design and maintain complex production data pipelines; develop efficient ETL/ELT processes; deploy changes through CI/CD; mentor junior engineers through code and design reviews; participate in problem management and root-cause analysis; partner with business and platform teams to expose curated datasets
Seniority
Senior, hands-on IC with mentorship responsibilities