9 дней назад
Engineering Manager, Platform (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Engineering Manager, Platform (AI) (Kubernetes/Go/Python): Leading platform engineering and observability teams building the infrastructure beneath Firmus AI Cloud, from bare-metal GPU compute and high-performance networking to self-service tooling and production services with an accent on reliability, security, developer experience, and cost efficiency. Focus on scaling teams and AI infrastructure, governing AI-assisted and agentic workloads, resolving L3 incidents, and delivering dependable platform services across expanding data centres.
Location: Singapore
Company
Technologies develops and operates energy-efficient AI infrastructure, including the AI Cloud GPU platform, liquid cooling systems, energy management, and AI software orchestration.
What you will do
- Lead the Platform Engineering and Observability team, including hiring, coaching, performance management, organisation design, and headcount planning.
- Own end-to-end delivery across multiple squads, coordinating roadmaps, dependencies, prioritisation, sprint planning, release gates, and delivery metrics.
- Set technical direction for bare-metal GPU compute, high-performance networking, internal platform services, self-service tooling, and observability products.
- Own production reliability, security, SLOs, on-call practices, change management, L3 incident resolution, SLA-breach response, and post-mortems.
- Govern the production path for AI-generated code and agentic workloads while extending SOC 2 Type 2 and ISO 27001 controls.
- Represent engineering to the CTO, product and architecture teams, delivery and operations, and enterprise or government customers.
Requirements
- Bachelor’s degree in computer science or a related technical field.
- 10+ years in platform or infrastructure engineering, including substantial hands-on experience as a senior or staff engineer.
- 4+ years managing engineers with direct ownership of hiring, development, and performance for a team of at least 8 engineers.
- Strong experience with Kubernetes, infrastructure-as-code, CI/CD, distributed systems, Go, and Python.
- Experience operating observability platforms at scale across metrics, logs, traces, and streaming pipelines.
- Experience with production reliability, security, compliance, cost management, and frameworks such as SOC 2 Type 2 or ISO 27001; clear written and verbal English is required.
Nice to have
- Experience running a platform as a product and measuring adoption, value, and developer experience.
- Experience with AI-assisted development and agentic AI workloads.
- Experience briefing enterprise or government customers on technical trade-offs.
Culture & Benefits
- Work at the intersection of sustainability, artificial intelligence, and next-generation infrastructure.
- Collaborate closely with founders and experienced technical leaders in an emerging company.
- Contribute to AI infrastructure designed to improve efficiency, scalability, and affordability.
- Participate in the incident-response on-call rotation and travel overseas occasionally when required.
- Work in an equal-opportunity environment that values diverse backgrounds and authentic collaboration.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 дней назад
Director, Software Engineer (AI)
13 дней назад
Engineering Manager - Cloud Platform & Operations
9 дней назад
VP of Engineering (SaaS)
14 дней назад
Senior Engineering Manager (Fraud)
14 дней назад
Analytics Engineering Manager (AI & Agentic Analytics)
11 дней назад
Thermal Engineering Leader (Physics AI)
1 - 2SGD