Lead Platform Engineer, Flink
Core
Design, evolve, and operate Grab's unified real-time data streaming platform using Apache Flink as the core compute engine to enable self-service capabilities for product and data teams.
Role type
Lead Platform Engineer (Stream Processing)
Builds
Self-serve streaming platform infrastructure, abstractions, modules, and libraries for Flink, Kafka, and real-time data pipelines.
Domain
Real-time data infrastructure, stream processing, distributed systems
Deliverable
production ML models | infrastructure
Required skills
Leading teams or projects in software/data/platform engineering, building and operating production stream processing pipelines (Apache Flink or Spark Streaming), hands-on engineering with Kafka and Scala/Java, distributed systems fundamentals, reliability engineering, production operations, technical design leadership, mentoring engineers.
Preferred skills
Kafka Connect, Kubernetes, Go, GitLab CI, AWS, Terraform, building reusable platform abstractions/SDKs, operating Apache Flink (HA, checkpointing, safe deployments), data warehouse/lake ecosystems (Spark, Parquet, Iceberg, Delta, Hudi).
Technologies
Apache Flink, Kafka, AWS, Kubernetes, Scala, Java, Go, Terraform, GitLab CI
Responsibilities
Design and build abstractions/modules/libraries to lower barriers to Flink adoption, improve automation and self-service workflows for scaling Flink pipelines, drive technical design discussions and production readiness reviews, partner with data ecosystem teams to ensure pipeline reliability, mentor engineers through design reviews and operational best practices.
Seniority
Senior, hands-on IC with leadership responsibilities
