Principal Software Engineer- Data Platforms
Core
Building, operating, and evolving reliable, scalable data platforms that support AI, analytics, and enterprise workflows for life sciences, diagnostics, and biotechnology.
Role type
Principal Software Engineer (Data Platforms)
Builds
Production data pipelines, datasets, and ingestion capabilities for structured, semi-structured, and unstructured data.
Domain
Life sciences, diagnostics, biotechnology, data engineering, AI/ML.
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Building and operating production data pipelines, handling structured and unstructured data (ingestion, transformation, enrichment, indexing), ensuring data quality and reliability, enabling data for downstream AI and reporting, technical leadership and mentorship.
Preferred skills
Supporting GenAI, NLP, or AI-driven discovery use cases, experience with cloud-based data platforms (Databricks, AWS), working in regulated or quality-sensitive domains (GxP-aligned).
Technologies
Databricks, AWS
Responsibilities
Lead hands-on delivery of scalable data pipelines powering AI and analytics; build and evolve data ingestion and processing capabilities; implement metadata and data quality practices; partner with architects and domain experts to ensure data usability; act as a technical leader setting standards and mentoring engineers.
Seniority
Principal, hands-on IC with mentorship