Research Data Specialist
Core
Lead and execute the management, integration, and architecture of biological data systems and computational infrastructure to support crop biotechnology research.
Role type
Senior IC data infrastructure engineer (bioinformatics)
Builds
Scalable storage solutions, HPC cluster environments, and automated data analysis pipelines for genomic and multi-omics datasets.
Domain
Agriculture biotechnology / Bioinformatics
Deliverable
production ML models | infrastructure
Required skills
HPC cluster management, cloud data warehousing, data pipeline architecture, Linux/Unix administration, SQL, Python, Bash/Shell scripting, version control, containerization
Preferred skills
Plant biology background, experience in agriculture or life sciences industry, knowledge of graph databases
Technologies
AWS (S3/EC2), Snowflake, Slurm, SGE, Linux/Unix
Responsibilities
Manage and optimize HPC cluster environments and cloud data warehouses; curate and version biological experimental data and bioinformatics software; partner with research teams to streamline data ingestion and automated analysis workflows; provide training on data submission protocols and cloud usage; evaluate emerging cloud storage and distributed computing frameworks to modernize infrastructure.
Seniority
Senior, hands-on IC