Research Data Engineer I
Core
Supports the implementation and maintenance of analytical and data science-based software and data pipelines enabling scientific workflows in a research enterprise.
Role type
Junior IC research data engineer
Builds
ETL data pipelines, data warehouses, data lakes, and custom research project-specific data workflow solutions
Domain
Biomedical research / Cancer Center / Genomics
Deliverable
production ML models | infrastructure
Required skills
SQL, Java, Python, R, Linux, Docker, AWS, Microsoft Azure, Google Cloud Platform, version control, software release management
Preferred skills
Genomics, metagenomics, flow cytometry, imaging files, metadata, data standards, laboratory data management systems (biospecimen software, electronic laboratory notebooks, LabKey Server, REDCap)
Responsibilities
Maintain and modify ETL data pipelines and overall data architecture; Support conversion of business and technical requirements into professional software solutions; Support implementation of custom research project-specific data workflow solutions; Participate in specification, implementation and execution of testing procedures; Produce and maintain comprehensive technical documentation
Seniority
Junior, hands-on IC