обновлено 1 месяц назад
DevOps Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
DevOps Engineer (AI/Linux): Building and administering automated Linux infrastructure, CI/CD pipelines, observability systems, and customer dashboards for high-performance computing and AI environments with an accent on infrastructure as code, systems integration, and performance troubleshooting. Focus on designing lights-out operations, resolving AI task and Linux environment issues, and supporting complex HPC, cloud, and enterprise systems through an on-call rotation.
Location: Hybrid in South Korea, with 3 days in the office. Office location: Seongnam-si, Gyeonggi-do, KR 13529.
Company
builds and manages AI factory platforms, high-performance computing infrastructure, memory and compute systems, infrastructure software, and end-to-end technology services.
What you will do
- Build automation workflows for lights-out operations.
- Integrate systems and platforms using infrastructure as code.
- Build and maintain CI/CD pipelines.
- Implement monitoring and observability for system performance metrics, alerting, and customer-service dashboards.
- Provide customer-facing IT support, troubleshoot Linux and AI task errors, and collaborate with customers on technical initiatives.
- Respond to after-hours system errors and participate in a weekly on-call rotation.
Requirements
- Bachelor’s degree in computer science, computer or electrical engineering, a related field, or equivalent experience.
- 6+ years of hands-on experience with UNIX/Linux server environments, CI/CD pipelines, and infrastructure as code.
- 2+ years of experience with Python and Ansible scripting.
- Strong Linux systems administration skills, including networking protocols, open-source technologies, performance troubleshooting, and system monitoring.
- Familiarity with CPU and GPU compilers and technologies including GCC, Intel, AMD AOCC/ROCm, NVIDIA PGI/OpenACC, and CUDA.
- Ability to speak English and communicate clearly with team members and clients.
Nice to have
- HPC application and systems management, MPI, optimization, hardware, and data center experience.
- AI virtualization, applications, container orchestration, systems management, or hardware design experience.
- High-performance storage and parallel file system experience, including HPC cluster administration.
- Linux certifications such as RHCSA or RHCE.
- Experience installing, configuring, tuning, and supporting software applications, including vendor coordination.
Culture & Benefits
- Collaborative environment focused on innovation, ownership, and servant leadership.
- Opportunities to work with advanced AI tools and technologies.
- Exposure to challenging infrastructure, HPC, cloud, and enterprise systems.
- Flexible, outcome-focused work environment with continuous learning opportunities.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
Cloud DevOps Engineer (AI)
9 дней назад
Senior DevOps Engineer (AI Security)
12 дней назад
DevOps Engineer (AWS)
100 000 - 145 000$
12 дней назад
DevOps Engineer (AI)
90 000 - 130 000$
TensorWave
12 дней назад
Senior Solutions Engineer (AI/GPU Infrastructure)
9 дней назад