Lead Data Engineer, GrabX (Experimentation & Config Management)
Core
Design and build high-throughput data pipelines to feed automated analysis engines for A/B testing and experiment management at Grab's superapp platform.
Role type
Lead Data Engineer (Experimentation & Config Management)
Builds
Real-time and batch data pipelines processing billions of experiment events daily for automated statistical analysis.
Domain
Ride-hailing and food delivery superapp; large-scale data engineering and experimentation infrastructure.
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data pipeline design (batch/real-time), Apache Spark, Apache Flink, OLAP databases (StarRocks/Trino/ClickHouse), distributed systems at petabyte scale, data modeling, team leadership, system design for low-latency environments.
Preferred skills
Internal experimentation platform experience, CI/CD tools, AWS, Terraform, Kubernetes, statistical concepts (p-values, confidence intervals).
Technologies
Apache Flink, Apache Spark, Trino, StarRocks, Delta Lake, Golang, AWS, Redis, ScyllaDB, DynamoDB, Kubernetes, Terraform.
Responsibilities
Design robust data pipelines for experiment events; develop efficient Spark/Flink jobs for metrics and significance; implement data validation and monitoring; lead transition to centralized Metric Store; mentor engineers; collaborate with Data Scientists on hypothesis testing automation.
Seniority
Senior, hands-on IC with leadership responsibilities