Junior Big Data Engineer
Core
Design, develop, and maintain scalable data pipelines and reporting-ready data assets within Palantir Foundry to support downstream analytics and business reporting.
Role type
Junior Big Data Engineer
Builds
Scalable data pipelines, data models, and reporting datasets for analytics
Domain
Big Data Engineering, Cloud Data Platforms
Deliverable
production ML models | product features
Required skills
Python, PySpark, Big Data technologies (Hadoop, Spark, Kafka, BigQuery), Cloud services (AWS Glue, Azure Data Factory, Google Cloud Dataflow), Data modeling, ETL/ELT, Database systems (PostgreSQL, MySQL, NoSQL), Containerization (Docker, Kubernetes)
Preferred skills
Market research data experience, Veeva CRM, Reltio, SAP, Palantir Foundry, Pharma/Healthcare background, Automation/pipeline optimization
Technologies
Palantir Foundry, Python, PySpark, Apache Hadoop, Spark, Kafka, BigQuery, AWS Glue, Azure Data Factory, Google Cloud Dataflow, PostgreSQL, MySQL, NoSQL, Docker, Kubernetes
Responsibilities
Design and implement scalable data ingestion, transformation, and orchestration pipelines; Develop reporting-ready datasets and data models; Optimize ETL/ELT processes for data integration; Implement data quality validation and monitoring; Monitor pipeline performance and troubleshoot production issues; Maintain technical documentation and promote automation
Seniority
Junior, hands-on IC