Principal Data Engineer
Core
Senior individual contributor leading cross-domain data engineering initiatives, defining enterprise standards, and delivering complex data architectures for analytics, machine learning, and AI.
Role type
Principal Data Engineer (IC)
Builds
Enterprise data capabilities including semantic/metrics layers, canonical data models, batch/API/event-driven/streaming patterns, and governed data products.
Domain
Healthcare diagnostics and oncology (Life Sciences)
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Spark on Databricks, Python, Scala, SQL, Snowflake, ETL/ELT pipeline design, relational/NoSQL data modeling, cloud infrastructure (AWS), REST API development, Agile tools
Preferred skills
Kafka, change data capture, event-streaming architectures, semantic layer design, multi-cloud data architecture (Azure/GCP), HIPAA/CLIA regulatory experience
Technologies
Databricks, Unity Catalog, Delta Lake, Apache Spark, Snowflake, AWS, S3, SQS, GitLab CI/CD, Tableau, JIRA, Confluence, Kafka
Responsibilities
Convert enterprise needs into executable architecture and technical plans; define and steward enterprise data engineering standards and reference architectures; lead enterprise design/code reviews and mentor Staff/Senior Engineers; evaluate architecture tradeoffs for reliability, scalability, and security; serve as technical escalation for multi-domain incidents; design architectures handling protected health information.
Seniority
Principal, hands-on IC with technical leadership