11 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Engineering Manager II, Data Platform (Apache Flink)
ΠΡΡΡ & Π‘ΠΎΠΏΡΠΎΠ²ΠΎΠ΄
ΠΠ»Ρ ΠΌΡΡΡΠ° Ρ ΡΡΠΎΠΉ Π²Π°ΠΊΠ°Π½ΡΠΈΠ΅ΠΉ Π½ΡΠΆΠ΅Π½ Plus
ΠΠΏΠΈΡΠ°Π½ΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
Π’Π΅ΠΊΡΡ:
TL;DR
Engineering Manager II, Data Platform (Apache Flink): Building and operating self-service stream processing infrastructure for real-time data pipelines with an accent on Apache Flink, deployment safety, reliability, and platform automation. Focus on leading the Vietnam-based engineering team, designing production guardrails, improving incident readiness, and scaling reusable capabilities across Grab's data ecosystem.
Location: Onsite at Vietnam office in Ho Chi Minh City
Company
is a Southeast Asian superapp providing transportation, food delivery, financial, and other digital services.
What you will do
- Lead a Vietnam-based team of Flink platform engineers, including hiring, onboarding, coaching, performance management, and career development.
- Own quarterly planning, objectives, prioritization, and delivery of the Flink platform roadmap.
- Guide technical design for self-service Flink capabilities, deployment guardrails, observability, automation, and reusable platform modules.
- Drive operational excellence through on-call readiness, incident response, post-incident learning, service level objectives, capacity planning, and reliability improvements.
- Partner with Kafka, data lake, metrics, governance, SRE, security, and product engineering teams across .
- Translate ambiguous platform needs into practical execution plans for technical and non-technical stakeholders.
Requirements
- 8+ years of professional software engineering, data engineering, or platform engineering experience, including 3+ years managing engineers or leading engineering teams.
- Experience building, operating, or managing distributed systems, data platforms, stream processing platforms, or high-scale backend infrastructure in production.
- Hands-on experience with Apache Flink, Spark Streaming, Kafka, or comparable large-scale data processing technologies.
- Strong knowledge of reliability engineering, production operations, observability, incident response, capacity planning, and security practices.
- Ability to set roadmaps, prioritize under ambiguity, lead execution, and deliver measurable outcomes through a team.
- Strong coaching, hiring, stakeholder management, and technical communication skills.
Nice to have
- Experience with Apache Flink internals, high availability, checkpointing, state management, job upgrades, shadow testing, and safe deployment patterns.
- Experience with Kubernetes, AWS, Terraform, GitLab CI, Go, Java, Scala, or infrastructure-as-code workflows.
- Experience building platform abstractions, SDKs, deployment tooling, control planes, self-service workflows, or managed NoOps services.
- Experience with data lake or warehouse technologies such as Spark, Parquet, Iceberg, Delta, Hudi, Trino, or Presto.
- Exposure to data governance, data quality, stream contracts, platform cost management, or multi-tenant data infrastructure.
Culture & Benefits
- Term life insurance and comprehensive medical insurance.
- Flexible benefits package through Flex.
- Parental leave, birthday leave, and volunteering leave.
- Confidential employee assistance programme.
- FlexWork arrangements, including differentiated working hours.
- Inclusive and equitable workplace with equal-opportunity employment practices.
ΠΡΠ΄ΡΡΠ΅ ΠΎΡΡΠΎΡΠΎΠΆΠ½Ρ: Π΅ΡΠ»ΠΈ ΡΠ°Π±ΠΎΡΠΎΠ΄Π°ΡΠ΅Π»Ρ ΠΏΡΠΎΡΠΈΡ Π²ΠΎΠΉΡΠΈ Π² ΠΈΡ ΡΠΈΡΡΠ΅ΠΌΡ, ΠΈΡΠΏΠΎΠ»ΡΠ·ΡΡ iCloud/Google, ΠΏΡΠΈΡΠ»Π°ΡΡ ΠΊΠΎΠ΄/ΠΏΠ°ΡΠΎΠ»Ρ, Π·Π°ΠΏΡΡΡΠΈΡΡ ΠΊΠΎΠ΄/ΠΠ, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡΠ΅ ΡΡΠΎΠ³ΠΎ - ΡΡΠΎ ΠΌΠΎΡΠ΅Π½Π½ΠΈΠΊΠΈ. ΠΠ±ΡΠ·Π°ΡΠ΅Π»ΡΠ½ΠΎ ΠΆΠΌΠΈΡΠ΅ "ΠΠΎΠΆΠ°Π»ΠΎΠ²Π°ΡΡΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡΠΈΡΠ΅ Π² ΠΏΠΎΠ΄Π΄Π΅ΡΠΆΠΊΡ. ΠΠΎΠ΄ΡΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β