Data Engineer, AI Enablement
Core
Design, build, and operate scalable data pipelines and curated data products to provide trusted, AI-ready data foundations for R&D analytics, knowledge graphs, and machine learning use cases.
Role type
Senior IC data engineer (AI enablement)
Builds
Curated, reusable data products and reliable data foundations for R&D analytics and AI
Domain
Life sciences / R&D data infrastructure
Deliverable
production ML models | product features
Required skills
SQL, Python, ETL/ELT patterns, workflow orchestration (Airflow), data modeling, metadata management, data governance, data quality checks, distributed SQL, cloud infrastructure, data integration, vector search preparation, knowledge graph data preparation
Preferred skills
Graph databases, ontology-based data structures, semantic data, linked-data concepts, AWS-based platforms, Databricks, Spark, Snowflake, Neo4j, regulated data environments
Technologies
Airflow, SQL, Python, Databricks, Spark, Snowflake, Neo4j, AWS
Responsibilities
Design and operate production data pipelines and curated data products; prepare data for AI, RAG, and knowledge graph use cases; apply data quality, governance, and lineage practices; collaborate with data scientists and ML engineers to translate requirements; provide technical guidance to contracted engineers; monitor pipeline performance and resolve operational issues.
Seniority
Senior, hands-on IC