13 дней назад
Senior Infrastructure Engineer (Kubernetes/AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Infrastructure Engineer (Kubernetes/AWS): Building and operating a multi-tenant Kubernetes and AWS foundation for distributed data compute engines with an accent on infrastructure automation, security, observability, and cost efficiency. Focus on designing custom Kubernetes operators, scaling compute platforms, implementing SRE practices, and managing reliable infrastructure changes across EKS, Terraform, and GitLab CI/CD.
Location: Onsite at the Petaling Jaya office, Malaysia
Company
is a Southeast Asian superapp providing food delivery, grocery, mobility, financial, and other digital services.
What you will do
- Design, build, and operate a multi-tenant Kubernetes EKS platform running Spark, Ray, Trino, Starrocks, Airflow, and Michelangelo workloads.
- Build Kubernetes operators and custom resources in Go to automate compute engine and tenant provisioning and lifecycle management.
- Lead Terraform infrastructure-as-code and GitLab CI/CD for AWS services including EKS, S3, IAM, RDS, VPC, and networking.
- Design identity, access, and security models across AWS IAM, Kubernetes RBAC, and service identities.
- Build observability, alerting, capacity planning, and incident tooling while contributing to SLOs, runbooks, on-call operations, and post-incident reviews.
- Improve compute cost efficiency and lead platform migrations, upgrades, cluster consolidation, and phased rollouts with rollback plans.
Requirements
- Degree in Software Engineering, Computer Science, or a related field.
- 3+ years of experience, including at least 2 years building and operating production infrastructure or platform services at scale.
- Programming proficiency in Go and/or Python, with experience treating infrastructure as tested, reviewed, versioned, and automated software.
- Deep production Kubernetes experience covering cluster operations, scheduling, autoscaling, networking, storage, RBAC, and multi-tenancy.
- Experience with AWS, including EKS, EC2, S3, IAM, and VPC, plus Terraform-based infrastructure as code.
- Proficiency with CI/CD, GitOps-style delivery, observability, SLOs, incident management, and capacity planning.
Nice to have
- Experience with Spark, Ray, Airflow, Trino, or Starrocks on Kubernetes.
- Knowledge of Kubernetes autoscaling and scheduling tools such as Karpenter, Cluster Autoscaler, Yunikorn, or Volcano.
- Experience with container runtimes, service mesh, ingress, network policies, and multi-tenant cluster operations.
- FinOps experience, including cost attribution and spot or reserved capacity strategies.
- Contributions to open-source cloud-native or data infrastructure projects.
Culture & Benefits
- Work in a data infrastructure team supporting company-wide data lake access and thousands of daily pipelines and queries.
- Term life insurance and comprehensive medical insurance.
- Flexible benefits through Flex, plus parental, birthday, and volunteering leave.
- Confidential employee assistance support and FlexWork arrangements with differentiated working hours.
- Inclusive workplace with equal-opportunity employment practices.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →