2 дня назад
Command Center Supervisor (HPC)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Command Center Supervisor (HPC): Leading 24/7 command center operations across data center environments with an accent on uptime, incident response, operational continuity, and team supervision. Focus on coordinating critical events, managing escalations and staffing, strengthening SOPs, and improving operational readiness across HPC infrastructure.
Location: On-site in Austin, Texas, United States; occasional travel between data centers. The role operates on nights, with overtime, weekends, and rotating on-call coverage as needed.
Company
provides high-performance computing infrastructure for AI, HPC, and digital asset mining operations across data center campuses in North America.
What you will do
- Lead real-time 24/7 command center operations, maintaining uptime, operational continuity, and effective incident response.
- Supervise Command Center Specialists, manage shift workloads and staffing coverage, and provide coaching and performance development.
- Coordinate cross-functional responses to security incidents, emergencies, crises, investigations, maintenance activities, and facility access operations.
- Serve as the escalation authority during critical events and participate in a rotating on-call schedule.
- Support business continuity and disaster recovery activities during planned and unplanned events.
- Improve workflows, SOPs, reporting quality, documentation accuracy, and process compliance across command center operations.
Requirements
- Associate degree in a related field or five years of relevant data center support experience.
- Previous leadership, supervisory, or team lead experience in a 24/7 operational environment.
- Strong understanding of data center infrastructure, HPC operational requirements, critical facility dependencies, and site configurations.
- Experience facilitating major incident response, escalation management, and cross-functional coordination.
- Strong analytical, organizational, decision-making, and communication skills for high-pressure operational events.
- Ability to work nights, weekends, overtime, rotating on-call shifts, and in data center environments that may include loud noise and construction.
Nice to have
- Understanding of ITIL service delivery concepts.
- Experience with Slack, Ignition, Metasys, CMMS systems, and Confluence.
- Experience with employee coaching, performance management, development planning, documentation audits, and ticket-quality reviews.
Culture & Benefits
- Entrepreneurial, collaborative, and results-driven operating culture.
- Opportunity to work with infrastructure supporting AI, HPC, and other high-compute workloads.
- Full-time schedule of 40 hours per week, Monday through Friday nights, with additional coverage requirements as needed.
- Work includes occasional travel between data center sites.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
Senior HPC Engineer (Classified Computing)
5 дней назад
Senior HPC Storage Engineer (HPC Storage)
6 дней назад
Primary Technical Control Facility Lead
7 дней назад
Digital Operations Center Leader (AI)
6 дней назад
Principal Computational Engineer (HPC)
114 400 - 216 320$
6 дней назад