Senior Data Engineer, GrabX
Core
Design and build high-throughput data pipelines to feed automated analysis engines for Grab's mission-critical Configuration Management and Experimentation platform.
Role type
Lead Data Engineer (Experimentation Platform)
Builds
High-throughput data pipelines and a centralized Metric Store for automated experiment analysis.
Domain
Ride-hailing and superapp ecosystem; large-scale data processing and experimentation.
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data pipeline design (batch/real-time), Apache Spark, Apache Flink, OLAP databases (StarRocks/ClickHouse/Trino), distributed systems architecture, data modeling, team leadership, system design.
Preferred skills
Internal experimentation platform experience, CI/CD tools, AWS, Terraform, Kubernetes, statistical concepts (p-values, confidence intervals).
Technologies
Apache Flink, Apache Spark, Trino, StarRocks, Delta Lake, Golang, AWS, Redis, ScyllaDB, DynamoDB, Kubernetes, Terraform.
Responsibilities
Design and build robust batch and real-time data pipelines; Develop efficient Spark or Flink jobs for complex metrics; Implement data validation and monitoring frameworks; Lead transition to a centralized Metric Store; Mentor engineers on data modeling and system design; Collaborate with Data Scientists on hypothesis testing automation.
Seniority
Senior, hands-on IC with leadership responsibilities