6 дней назад
Cloud Systems Engineer (AI/HPC)
100 000 - 135 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Cloud Systems Engineer (AI/HPC): Operating and scaling GPU-accelerated compute infrastructure for AI, machine learning, and data-intensive workloads with an accent on Linux administration, hardware lifecycle management, and datacenter operations. Focus on diagnosing enterprise compute failures, optimizing system performance, supporting InfiniBand and Ethernet environments, and building reliable operational automation.
Location: Tysons, Virginia, United States; employees work from the office 4 days a week
Base salary: $100,000–$135,000 USD per year
Company
provides an intelligently connected property platform covering security, video surveillance, access control, automation, energy management, and wellness solutions.
What you will do
- Deploy, configure, maintain, and operate GPU-accelerated compute infrastructure for AI training, inference, and data processing.
- Manage Linux operating systems, firmware, BIOS, BMC, drivers, patching, hardening, and lifecycle updates.
- Monitor infrastructure health, performance, utilization, capacity, and observability metrics.
- Install and troubleshoot enterprise hardware, including GPUs, CPUs, memory, storage, power, and networking components.
- Support rack-and-stack deployments, datacenter expansions, technology refreshes, hardware replacements, and vendor engagements.
- Collaborate with networking, storage, cloud, and AI engineering teams on scalability, resiliency, performance optimization, and incident response.
Requirements
- Bachelor’s degree.
- 3–5 years of production Linux systems administration and enterprise server infrastructure experience.
- Experience with large-scale compute environments, HPC platforms, AI infrastructure, or GPU-enabled systems.
- Experience with hardware diagnostics, firmware management, lifecycle maintenance, and datacenter operations.
- Knowledge of Bash, Python, PowerShell, operating system performance tuning, monitoring, storage, and networking fundamentals.
- New applicants requiring employment authorization sponsorship or other immigration-related support are not eligible for sponsorship for this position.
Culture & Benefits
- Collaborative, casual work environment with in-person teamwork.
- Medical plans with company subsidies and a company-contributed Health Savings Account.
- 401(k) with employer match.
- Paid vacation, holidays, wellness time, and maternity and bonding leave.
- Company-paid disability and life insurance.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
7 дней назад
High Performance Computing (HPC) AI Engineer
101 494 - 140 000$
9 дней назад
HPC Windows Server Infrastructure Engineer (AI)
136 300 - 199 900$
12 дней назад
HPC AI Systems Administrator Intern (AI)
35 - 41$
Nebius
6 дней назад
Senior HPC Cluster Engineer (AI)
8 дней назад
Datacenter Operations Engineer (AI Infrastructure)
150 000 - 300 000$
8 дней назад
System Administrator (Linux)
84 975 - 147 755$