Principal Data Engineer (Data Platforms)
Core
Design and build scalable, secure, cloud-native and hybrid data platforms and control-plane services to manage large-scale data storage, governance, and access across public-cloud and on-premises environments.
Role type
Principal Data Engineer (Data Platforms)
Builds
Scalable distributed and lakehouse platforms, control-plane services, and federated data-sharing solutions.
Domain
Payments industry, Cloud Data Engineering, Distributed Systems
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Python, PySpark, Cloud Architecture (AWS/Azure), Kubernetes, Distributed Systems Design, Data Governance, API Design (REST/gRPC), Data Modeling, CI/CD
Preferred skills
Lakehouse patterns (Iceberg, Delta Lake), Federated Metadata, Data Contracts, Sovereign Cloud Environments
Technologies
Java, Python, PySpark, AWS (S3, IAM, EKS, Glue), Azure (ADLS, Entra ID, AKS), Databricks, Snowflake, Spark, Trino, Iceberg, Delta Lake, Parquet, Unity Catalog, Apache Polaris, Kubernetes, gRPC
Responsibilities
Design scalable distributed and lakehouse platforms spanning regions and execution environments; Define control-plane and data-plane responsibilities and boundaries; Design federated metadata, catalog, identity, access, and query patterns; Create architecture decision records, reference architectures, and engineering standards; Mentor engineers and foster technical growth; Lead architectural discussions and influence technical direction.
Seniority
Principal, hands-on IC with mentorship