Middle/Senior Big Data Engineer
Core
Design, develop, and maintain scalable data pipelines and reporting-ready data assets within Palantir Foundry for a global healthcare organization.
Role type
Senior Big Data Engineer
Builds
Scalable data pipelines, data models, and reporting datasets for pharmaceutical and biotechnology analytics.
Domain
Healthcare / Life Sciences / Big Data Engineering
Deliverable
production ML models | product features
Required skills
Python, PySpark, Apache Spark, Apache Kafka, AWS Glue, Azure Data Factory, Google Cloud Dataflow, data modeling, ETL/ELT, PostgreSQL, MySQL, NoSQL, Docker, Kubernetes
Preferred skills
Palantir Foundry, Veeva CRM, Reltio, SAP, market research data, automation
Technologies
Palantir Foundry, Python, PySpark, Apache Hadoop, Apache Spark, Apache Kafka, BigQuery, AWS Glue, Azure Data Factory, Google Cloud Dataflow, PostgreSQL, MySQL, NoSQL, Docker, Kubernetes
Responsibilities
Design and implement scalable data ingestion, transformation, and orchestration pipelines; Develop reporting-ready datasets to support Power BI reporting; Optimize ETL/ELT processes for timely data delivery; Implement data quality validation and monitoring; Troubleshoot production issues and improve pipeline scalability; Maintain technical documentation and promote automation.
Seniority
Senior, hands-on IC