обновлено 1 месяц назад
Sr. DevOps Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Sr. DevOps Engineer (Linux/AI infrastructure): Building automation workflows, infrastructure-as-code integrations, CI/CD pipelines, and observability dashboards for complex Linux, HPC, cloud, and enterprise environments with an accent on systems administration, performance troubleshooting, and customer-facing operations. Focus on supporting AI workloads, integrating CPU and GPU compiler ecosystems, resolving production issues, and participating in on-call operations.
Location: Hybrid in South Korea, with 3 days in the office in Seongnam-si, Gyeonggi-do
Company
is an AI technology and infrastructure company that designs, builds, deploys, and manages high-performance computing, memory, software, and enterprise solutions.
What you will do
- Build automation workflows supporting lights-out operations.
- Integrate systems and platforms using infrastructure as code.
- Build and maintain CI/CD pipelines.
- Implement monitoring, observability, performance metrics, alerting, and customer service dashboards.
- Provide Linux systems support, troubleshoot errors including AI task errors, and collaborate with customers on technical initiatives.
- Respond to system errors after hours and participate in a weekly on-call rotation.
Requirements
- Bachelor’s degree in Computer Science, Computer or Electrical Engineering, a related field, or equivalent experience.
- 8+ years of hands-on experience with UNIX/Linux server environments, CI/CD pipelines, and infrastructure as code.
- 5+ years of experience with Python and Ansible scripting.
- Strong Linux systems administration, open-source technology, networking, monitoring, and performance troubleshooting skills.
- Familiarity with CPU and GPU compiler ecosystems, including GCC, Intel, AMD AOCC/ROCm, and NVIDIA PGI OpenACC/CUDA.
- Ability to speak English and communicate clearly with team members and clients.
Nice to have
- Experience with HPC applications, systems management, MPI, optimization, hardware, and data center environments.
- AI virtualization, applications, container orchestration, systems management, or hardware design experience.
- High-performance storage, parallel file systems, HPC cluster administration, and Linux cluster optimization experience.
- Linux certifications such as RHCSA or RHCE.
- Experience installing, configuring, tuning, and supporting software applications, including collaboration with application vendors.
Culture & Benefits
- Work in a customer-facing managed services environment supporting complex integrated systems.
- Collaborate across global R&D, manufacturing, and sales operations.
- Work with advanced AI, HPC, cloud, enterprise, and data center technologies.
- Environment focused on ownership, innovation, continuous learning, and outcome-driven work.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →