Senior Software Engineer, Data Engineering
Core
Design, build, and maintain data ingestion, transformation, and orchestration pipelines to support digital validation and manufacturing intelligence for life sciences companies.
Role type
Senior IC data engineering engineer
Builds
Data Lakehouse architectures, ETL/ELT pipelines, ML-ready datasets, and production data APIs
Domain
Life Sciences / Data Engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data pipeline orchestration, Data Lakehouse architecture, ETL/ELT development, Data integration, Data quality & governance, CI/CD, API development
Preferred skills
Azure Synapse, Delta Lake, Apache Spark, Kubernetes, Azure ML
Technologies
Azure Data Factory, Databricks, Apache Spark, Azure Synapse, Delta Lake, Power BI, Superset, Tableau, Azure ML, Kubernetes
Responsibilities
Design, develop, and maintain data ingestion, transformation, and orchestration pipelines (batch and real-time). Build and optimize data Lakehouse architectures using Azure Synapse, Delta Lake, or similar frameworks. Integrate and manage structured and unstructured data sources (SQL/NoSQL, files, documents, IoT streams). Develop and operationalize ETL/ELT pipelines using Azure Data Factory, Databricks, or Apache Spark. Collaborate with Data Scientists to prepare and serve ML-ready datasets for model training and inference. Implement data quality, lineage, and governance frameworks across pipelines and storage layers. Work with BI tools (Power BI, Superset, Tableau) to enable self-service analytics for business teams. Deploy and maintain data APIs and ML models in production using Azure ML, Kubernetes, and CI/CD pipelines.