Sr Data Engineer
Core
Design, build, and maintain scalable data pipelines and ETL/ELT processes to ingest, transform, and process large-scale structured and unstructured data for actionable business insights.
Role type
Senior IC Data Engineer (Big Data)
Builds
End-to-end data pipelines, data models, and automated workflows on cloud platforms
Domain
Biotechnology / Big Data Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Big data technologies (Databricks, Apache Spark/PySpark), SQL, Python, ETL/ELT processes, data modeling, data governance, cloud platforms (AWS), workflow orchestration, performance tuning
Preferred skills
Data protection regulations (GDPR, CCPA), machine learning model development, Apache Airflow, multi-source integration (APIs, cloud storage), global team collaboration
Technologies
Databricks, PySpark, SparkSQL, AWS, SQL, Python, Apache Airflow, SageMaker
Responsibilities
Design and develop data pipelines leveraging Databricks and PySpark to ingest and process large-scale datasets; Engineer solutions for both structured and unstructured data; Implement automated workflows for data ingestion, transformation, and deployment; Apply performance optimization techniques for Spark jobs; Build integrations with SQL databases, APIs, and cloud storage platforms; Manage data pipeline projects from inception to deployment; Collaborate with Data Architects, Business SMEs, and Data Scientists to design end-to-end data solutions.
Seniority
Senior, hands-on IC
