Senior Java Engineer (Data Engineering) (CPT Hybrid)
Core
Design and implement large-scale data migration strategies, build data lineage mapping systems, and develop scalable batch and streaming data transformation solutions for petabyte-scale enterprise data.
Role type
Senior IC Data Engineer with technical leadership
Builds
Production data pipelines, migration frameworks, and observability solutions for enterprise data platforms
Domain
Data Engineering / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Large-scale data migration, ETL pipeline development, Java/Scala programming, Distributed systems design, Cloud data platforms (AWS/GCP/Azure/Databricks), Kubernetes (EKS), Data privacy and compliance
Preferred skills
Apache Spark/PySpark, Hadoop ecosystem, Data lineage analysis, Modern data formats (Iceberg/Parquet/ORC/Avro), Streaming platforms (Kafka/Kinesis), Observability tools (Prometheus/Grafana)
Technologies
Java, Scala, Gradle, Maven, HDFS, S3, GCS, AWS EKS, Apache Spark, Apache Kafka, Prometheus, Grafana
Responsibilities
Design large-scale data migration strategies for identity transformation, Build data lineage mapping and validation systems, Develop scalable data transformation solutions for batch and streaming processing, Implement monitoring and observability solutions for data pipelines, Create testing and validation frameworks for data accuracy, Provide technical leadership and mentor engineers on data pipeline development
Seniority
Senior, hands-on IC with technical leadership