Middle Big Data Engineer
Core
Design, develop, and maintain scalable data pipelines and reporting-ready data assets within Palantir Foundry to enable reporting and data transformation.
Role type
Middle Big Data Engineer
Builds
Scalable data pipelines, reporting datasets, and ETL/ELT processes for a global healthcare organization
Domain
Healthcare / Life Sciences / Big Data Engineering
Deliverable
production ML models | product features
Required skills
Python, PySpark, Big Data technologies (Hadoop, Spark, Kafka, BigQuery), Cloud services (AWS Glue, Azure Data Factory, Google Cloud Dataflow), Data modeling, Data warehousing, ETL/ELT, Database systems (PostgreSQL, MySQL, NoSQL), Containerization (Docker, Kubernetes)
Preferred skills
Market research data experience, Veeva CRM, Reltio, SAP, Palantir Foundry, Pharma/Healthcare background, Automation/pipeline optimization
Technologies
Palantir Foundry, Python, PySpark, Apache Hadoop, Spark, Kafka, BigQuery, AWS Glue, Azure Data Factory, Google Cloud Dataflow, PostgreSQL, MySQL, NoSQL, Docker, Kubernetes
Responsibilities
Design and implement scalable data pipelines in Palantir Foundry; Develop reporting-ready datasets and data models; Optimize ETL/ELT processes for data integration; Implement data quality validation and monitoring; Monitor pipeline performance and troubleshoot production issues; Maintain technical documentation and promote automation
Seniority
Middle, hands-on IC