Staff / Principal Data Engineer
Core
Own the unified data platform and data lake powering fraud detection and response, serving as the foundation for AI models and downstream capabilities.
Role type
Staff / Principal Data Engineer
Builds
Unified data platform, data lake, low-latency batch and streaming pipelines, and generative AI data infrastructure (embedding generation, vector storage).
Domain
Fraud protection, cybersecurity, threat intelligence, big data
Deliverable
production ML models
Required skills
Large-scale data platform architecture, Apache Spark, Apache Flink, batch and streaming pipeline engineering, Kubernetes, CI/CD, data quality and observability
Preferred skills
Generative AI and embedding models, vector databases, cybersecurity or threat intelligence background, transaction fraud signals
Technologies
Apache Spark, Apache Flink, Kubernetes, vector databases
Responsibilities
Design, build, and operate the data lake and ingestion platform end-to-end; build low-latency pipelines ingesting and normalizing internal/external signals; establish data quality, freshness, lineage, and observability; deploy and maintain platform reliability on Kubernetes; partner with data science and product teams to enable shared data foundations.
Seniority
Staff / Principal, hands-on IC with limited supervision