Data Engineer PySpark - Data Factory - Services Financiers - Ile de France
Core
Building and maintaining a Data Factory for a major state financial actor, focusing on data ingestion, DataLake structuration, governance, and processing from batch/streaming to business exposure.
Role type
Senior Data Engineer (PySpark/Cloud)
Builds
Data ingestion pipelines (batch/streaming), DataLake, CI/CD chains, and Proof of Concepts/MVPs for financial clients.
Domain
Financial Services / Big Data Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Python, SQL, Scala, R, Java, PySpark, Hadoop, Kafka, RabbitMQ, Cloud platforms, CI/CD, Data governance
Preferred skills
Data modeling, Agile/Scrum methodology, mentoring, technology scouting
Technologies
PySpark, Hadoop, Kafka, RabbitMQ, Cloud (AWS/Azure/GCP implied)
Responsibilities
Translate business requirements into data engineering solutions; implement batch and streaming data ingestion; structure DataLake and enforce governance; process data for business exposure; maintain CI/CD pipelines; develop prototypes and MVPs.
Seniority
Senior, hands-on IC