Translational Data Management, Automation, & AI Engineer
Core
Design, build, and operate robust biomarker and clinical data ingestion pipelines to feed a biomarker platform for analysis, visualization, and machine-learning use cases supporting clinical trials.
Role type
Senior IC data engineering and AI automation engineer
Builds
End-to-end data ingestion pipelines, automated validation workflows, and harmonized clinical datasets
Domain
Clinical research / Precision medicine / Bioinformatics
Deliverable
production ML models
Required skills
Python, database design, Databricks, workflow orchestration (Airflow, Nextflow, Snakemake), agentic automation, AI workflow development, HPC, cloud platforms (AWS), Git, CI/CD, Docker, clinical data standards (CDISC/SDTM/ADaM), R pipelines
Preferred skills
PhD, experience with biomarker/biological/clinical data, experience with clinical labs and CROs, experience with specific assay types (immunoassay, flow cytometry, proteomics, sequencing)
Technologies
Python, Databricks, Airflow, Nextflow, Snakemake, AWS, Git, Docker, R
Responsibilities
Design and maintain data ingestion pipelines; implement automated data validation and quality control; integrate agentic automation and generative AI; collaborate with labs and CROs to onboard assays; build harmonization logic and data models; generate study-specific analysis bundles; write production-grade Python code and contribute to CI/CD
Seniority
Senior, hands-on IC