Senior Data Engineer
Core
Design, build, and operate modern Databricks and Lakehouse data platforms supporting advanced analytics, AI, machine learning, and Generative AI initiatives.
Role type
Senior individual contributor data engineer (lakehouse & GenAI enablement)
Builds
Scalable data pipelines, lakehouse structures, curated knowledge datasets, and production-ready data solutions for AI workloads
Domain
Enterprise data engineering, cloud-native lakehouse architecture, Generative AI data patterns
Deliverable
production ML models | product features
Required skills
Python, SQL, Apache Spark, Databricks, Delta Lake, Unity Catalog, CI/CD, data governance, scalable pipeline design
Preferred skills
RAG data patterns, feature-store workflows, vector/embedding-ready data, AWS or Azure cloud platforms
Technologies
Databricks, Delta Lake, Unity Catalog, Python, PySpark, SQL, Apache Spark
Responsibilities
Design and maintain scalable data pipelines and lakehouse structures; Develop pipelines supporting Generative AI applications including curated datasets and RAG patterns; Implement automated testing and CI/CD practices for production-grade solutions; Collaborate with AI/ML Engineers and Product Owners on AI-focused delivery initiatives; Ensure data solutions are observable, resilient, performant, and cost-efficient.
Seniority
Senior, hands-on IC