Lead Data Engineer
Core
Lead the modernization and professionalization of data infrastructure, platforms, and pipelines for Flagship Pioneering's portfolio-facing IT organization.
Role type
Senior individual-contributor Lead Data Engineer
Builds
Cloud-native data architectures, data lakes, marts, warehouses, and complex integration pipelines for emerging biotech companies.
Domain
Biotechnology / Life Sciences / Cloud Data Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Cloud-native data architecture design, Python, SQL, Dagster, dbt, Spark, Iceberg, AWS (Lake Formation, Athena, Glue), Git, CI/CD, CDK, Docker, ECS, structured and unstructured data handling, generative AI tooling
Preferred skills
Bioinformatics/cheminformatics workflows, scientific software (Benchling, CDD), Agile/Scrum methodologies, AWS certifications
Technologies
AWS, Dagster, dbt, Spark, Iceberg, Lake Formation, Athena, Glue, Docker, ECS, CDK
Responsibilities
Lead implementation of data infrastructure and platforms; optimize data storage solutions (lakehouses, marts, warehouses); implement complex backend and data integration pipelines; establish engineering standards for testing, documentation, and monitoring; diagnose and resolve complex infrastructure and pipeline failures; mentor junior engineers; leverage generative AI tools to accelerate pipeline development.
Seniority
Senior, hands-on IC