5 дней назад
Senior Manager (AI Data Center Operations)
175 000 - 200 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Manager (AI Data Center Operations): Leading a 24/7/365 Facility Remote Operations Center monitoring electrical, mechanical, fire/life safety, and building automation systems across a national AI data center portfolio with an accent on MEP technical oversight, incident response, and operational team leadership. Focus on building scalable SOPs and runbooks, reducing detection and escalation times, and coordinating complex facility events across remote and on-site teams.
Location: On-site in Dallas, Texas, United States; supports a 24/7/365 operation and off-hours escalations.
Salary: $175,000–$200,000 per year plus bonus and restricted stock units.
Company
builds and operates AI-optimized data centers and energy infrastructure for high-density AI workloads.
What you will do
- Lead 24/7/365 Facility Remote Operations Center operations, including staffing, scheduling, shift coverage, and performance management.
- Build and improve monitoring protocols, escalation procedures, SOPs, emergency runbooks, and escalation matrices for facility alarms and events.
- Oversee monitoring of electrical distribution, backup power, mechanical cooling, fire/life safety, BMS, DCIM, and EPMS systems.
- Manage detection, notification, escalation, documentation, root cause analysis, and after-action reviews for critical facility incidents.
- Coordinate with Data Center Operations, Critical Facilities Engineering, Construction/Commissioning, Security, vendors, and leadership.
- Recruit, train, and develop ROC staff while improving automation, tooling, and alarm management processes.
Requirements
- 5+ years of data center operations experience with direct responsibility for critical facility uptime.
- Strong knowledge of MEP systems, electrical distribution, generators, UPS, and air- and liquid-cooled systems in high-density AI/HPC data centers.
- Experience working in shift-based 24/7 operations, including rotating schedules, staffing, and people management.
- Experience with incident management, escalation procedures, root cause analysis, and DCIM, BMS/EPMS, or similar monitoring platforms.
- Strong communication skills for translating technical facility issues into actionable information.
- Must be able to work on-site in Dallas, Texas and support a 24/7 operation, including off-hours escalations.
Nice to have
- Hyperscale, colocation, or AI/GPU-cluster data center operations experience.
- CDCP, CDCS, CDCE, DCPRO, or equivalent certification.
- Experience building or scaling a remote or centralized operations function.
- Experience with liquid cooling infrastructure and geographically distributed critical infrastructure portfolios.
- Bachelor’s degree in Electrical Engineering, Mechanical Engineering, Facilities Management, or a related technical field, or equivalent practical experience.
Culture & Benefits
- Competitive compensation, bonus, equity, and restricted stock units.
- Paid time off, holidays, leave programs, and parental leave.
- Health, dental, vision, life, disability, and mental health support.
- HSA contributions and a 401(k) plan with company match up to 4% of salary.
- Professional development, tuition reimbursement, commuter benefits, cell phone stipend, meals allowance, and travel insurance.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Engineering Technical Lead – Engineering Triage
170 000 - 210 000$
4 дня назад
Sr. Director of Enterprise Systems & Data (AI)
208 000 - 286 000$
4 дня назад
Head of Project Engineering (AI Infrastructure)
225 000 - 330 000$
4 дня назад
Engineering Manager, BMS
5 дней назад
Manager, Software Engineering (AI)
155 000 - 233 000$
4 дня назад
Senior Engineering Manager (AI)
215 000 - 250 000$