Data Engineer – Computational Biology (Senior Associate)
Core
Develop, operationalize, and evolve production bioinformatics pipelines and data products for Pfizer R&D research units to enable drug discovery.
Role type
Senior Associate Data Engineer (Computational Biology)
Builds
Production-grade Nextflow pipelines and omics data platforms for drug discovery
Domain
Pharmaceutical R&D / Computational Biology
Deliverable
production ML models
Required skills
Nextflow, Python, cloud infrastructure, workflow lifecycle management, DevOps practices, troubleshooting, data integration
Preferred skills
AI-assisted coding tools, software engineering best practices, large heterogeneous data set processing
Technologies
Nextflow, Python, cloud infrastructure
Responsibilities
Develop and deploy production-grade Nextflow pipelines on cloud infrastructure; Support pipeline lifecycle management including upgrades and performance tuning; Apply DevOps best practices for platform services; Translate data analysis requirements into robust pipeline solutions; Develop and evolve an omics data platform; Collaborate with external partners to strengthen pipeline quality
Seniority
Senior Associate, hands-on IC