Staff Data Engineer - Governance
Core
Design, implement, and optimize distributed data systems and real-time pipelines powering analytics, machine learning, and AI-driven experiences.
Role type
Staff Data/Platform Engineer
Builds
Scalable, low-latency data platforms, streaming pipelines, and ML feature stores
Domain
Data Engineering, Distributed Systems, Real-time Processing
Deliverable
production ML models | infrastructure
Required skills
Distributed systems design, Streaming architectures, Apache Kafka, Apache Flink, Apache Spark, Delta Lake, SQL, Data modeling, CI/CD, Software design patterns
Preferred skills
Model Context Protocol (MCP), ML Ops, Feature stores, Data governance, Observability stacks
Technologies
Kafka, Flink, Spark, Delta Lake, Scala, Java, Python
Responsibilities
Design and optimize distributed data systems for scalability and fault tolerance; Build and maintain low-latency pipelines handling billions of events; Translate business requirements into robust architectural patterns; Mentor engineers and establish code quality standards; Bridge traditional ML workloads with modern AI decision systems; Collaborate cross-functionally to drive platform adoption.
Seniority
Staff, hands-on IC with mentorship