What you'd do:
- Lead design and operation of Grab's Apache Flink stream processing platform
- Drive medium and large projects across Data Engineering Platforms organization
- Mentor senior engineers through design reviews, code reviews, and debugging sessions
- Build self-serve platform abstractions that make Flink adoption easier and safer
- Improve automation workflows supporting many production Flink pipelines at scale
- Partner with Kafka, data lake, metrics, and governance teams on reliable pipelines
- Lead technical design discussions, production readiness reviews, and incident learning
- Work hands-on across Flink, Kafka, AWS, Kubernetes, observability, and SRE practices
What they want:
- Five or more years leading software, data, or platform engineering projects
- Production experience building stream processing pipelines, preferably Flink or Spark Streaming
- Strong hands-on Kafka experience with Scala or Java programming skills
- Solid distributed systems, scalable processing, reliability, and production operations fundamentals
- Ability to lead technical design, mentor engineers, and drive projects to production
- Willingness to learn new technologies and improve platform reliability for internal users
Nice to have:
- Experience with Kafka Connect, Kubernetes, Go, GitLab CI, AWS, or Terraform
- Experience building platform abstractions, SDKs, deployment tooling, or self-service workflows
- Production Apache Flink operations including HA, checkpointing, safe deployments, incident response
- Familiarity with data lake ecosystems such as Spark, Parquet, Iceberg, Delta, or Hudi Five or more years leading software, data, or platform engineering projects
- Production experience building stream processing pipelines, preferably Flink or Spark Streaming
- Strong hands-on Kafka experience with Scala or Java programming skills
- Solid distributed systems, scalable processing, reliability, and production operations fundamentals
- Ability to lead technical design, mentor engineers, and drive projects to production
- Willingness to learn new technologies and improve platform reliability for internal users
Production experience building stream processing pipelines, preferably Flink or Spark Streaming
Strong hands-on Kafka experience with Scala or Java programming skills
Solid distributed systems, scalable processing, reliability, and production operations fundamentals
Ability to lead technical design, mentor engineers, and drive projects to production
Willingness to learn new technologies and improve platform reliability for internal users