2 ΠΌΠ΅ΡΡΡΠ° Π½Π°Π·Π°Π΄
Staff/Senior Staff Engineer, Kubernetes (Web3)
ΠΡΡΡ & Π‘ΠΎΠΏΡΠΎΠ²ΠΎΠ΄
ΠΠ»Ρ ΠΌΡΡΡΠ° Ρ ΡΡΠΎΠΉ Π²Π°ΠΊΠ°Π½ΡΠΈΠ΅ΠΉ Π½ΡΠΆΠ΅Π½ Plus
ΠΠΏΠΈΡΠ°Π½ΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
Π’Π΅ΠΊΡΡ:
TL;DR
Staff/Senior Staff Engineer, Kubernetes (Web3): Managing large-scale production Kubernetes clusters and multi-cloud infrastructure on AWS and Alibaba Cloud with an accent on high availability, performance tuning, and disaster recovery. Focus on building automated DevOps pipelines, strengthening cluster security, and leading incident response for a global crypto exchange.
Location: Must be based in Singapore with a current right to work (no visa sponsorship provided)
Company
OKX is a leading global crypto exchange and developer of decentralized applications, committed to fostering financial freedom through blockchain technology.
What you will do
- Own the lifecycle management, scaling, and performance tuning of large-scale production Kubernetes clusters.
- Govern and optimize multi-cloud resources across AWS and Alibaba Cloud environments.
- Lead containerization and microservices operational rollouts while resolving complex scheduling and network challenges.
- Build and maintain comprehensive monitoring, alerting, and distributed tracing systems.
- Develop automation scripts and integrate CI/CD pipelines using IaC principles to improve operational efficiency.
- Lead incident response, conduct root cause analysis, and implement long-term stability improvements.
Requirements
- Must have current right to work in Singapore (no sponsorship).
- Bachelor's degree or higher in a computer-related field.
- 4+ years of hands-on experience operating production-level Kubernetes clusters.
- Proficiency in Alibaba Cloud and AWS dual-cloud operations and disaster recovery.
- Strong automation skills in Shell and Python, with experience in CI/CD tools like Jenkins, GitLab CI, and ArgoCD.
- Deep understanding of Linux system administration, container runtimes, and observability stacks.
Nice to have
- Experience in large-scale public cloud environments (100+ nodes).
- Kubernetes CKA/CKS certification.
- Experience with AI/LLM workload scheduling and GPU management.
- Knowledge of K8s security hardening tools like OPA/Gatekeeper or Falco.
Culture & Benefits
- Competitive total compensation package.
- Comprehensive healthcare schemes for employees and dependants.
- L&D programs and education subsidies for professional growth.
- Wellness and meal allowances.
- Diverse, friendly, and rewarding work environment.
ΠΡΠ΄ΡΡΠ΅ ΠΎΡΡΠΎΡΠΎΠΆΠ½Ρ: Π΅ΡΠ»ΠΈ ΡΠ°Π±ΠΎΡΠΎΠ΄Π°ΡΠ΅Π»Ρ ΠΏΡΠΎΡΠΈΡ Π²ΠΎΠΉΡΠΈ Π² ΠΈΡ ΡΠΈΡΡΠ΅ΠΌΡ, ΠΈΡΠΏΠΎΠ»ΡΠ·ΡΡ iCloud/Google, ΠΏΡΠΈΡΠ»Π°ΡΡ ΠΊΠΎΠ΄/ΠΏΠ°ΡΠΎΠ»Ρ, Π·Π°ΠΏΡΡΡΠΈΡΡ ΠΊΠΎΠ΄/ΠΠ, Π½Π΅ Π΄Π΅Π»Π°ΠΉΡΠ΅ ΡΡΠΎΠ³ΠΎ - ΡΡΠΎ ΠΌΠΎΡΠ΅Π½Π½ΠΈΠΊΠΈ. ΠΠ±ΡΠ·Π°ΡΠ΅Π»ΡΠ½ΠΎ ΠΆΠΌΠΈΡΠ΅ "ΠΠΎΠΆΠ°Π»ΠΎΠ²Π°ΡΡΡΡ" ΠΈΠ»ΠΈ ΠΏΠΈΡΠΈΡΠ΅ Π² ΠΏΠΎΠ΄Π΄Π΅ΡΠΆΠΊΡ. ΠΠΎΠ΄ΡΠΎΠ±Π½Π΅Π΅ Π² Π³Π°ΠΉΠ΄Π΅ β
ΠΠΎΡ ΠΎΠΆΠΈΠ΅ Π²Π°ΠΊΠ°Π½ΡΠΈΠΈ
5 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Site Reliability Engineer - Vice President
3 Π΄Π½Ρ Π½Π°Π·Π°Π΄
DataOps Engineer (AWS/Kubernetes)
4 Π΄Π½Ρ Π½Π°Π·Π°Π΄
Cloud Senior DevOps Engineer
Latitude
3 Π΄Π½Ρ Π½Π°Π·Π°Π΄
Senior Site Reliability Engineer (Kubernetes)
7 Π΄Π½Π΅ΠΉ Π½Π°Π·Π°Π΄
Senior DevOps Specialist (Web3)
165Β 000 - 195Β 000$
IG
3 Π΄Π½Ρ Π½Π°Π·Π°Π΄