обновлено 6 дней назад
Network Engineer (AI/HPC)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Network Engineer (AI/HPC): Managing and improving InfiniBand and Ethernet networks in large-scale Linux-based HPC and AI data centers with an accent on network reliability, diagnostics, automation, and performance. Focus on optimizing low-latency, high-bandwidth infrastructure, resolving complex network issues, and supporting VXLAN, EVPN, BGP, and OSPF environments.
Location: Hybrid in South Korea, with 3 days in the office in Seongnam-si, Gyeonggi-do
Company
is an end-to-end technology company building and managing AI factory platforms, HPC infrastructure, advanced memory, compute systems, and related enterprise solutions.
What you will do
- Monitor, maintain, upgrade, and manage InfiniBand and Ethernet network equipment.
- Diagnose and resolve network and server issues across large-scale data center and HPC environments.
- Run hardware diagnostics and replace failing components.
- Support Linux-based high-performance computing and AI infrastructure, monitoring network processes and data flow.
- Collaborate with software and network engineering teams on cybersecurity and network efficiency.
- Develop operational automation, maintain technical documentation, and supervise on-site component updates.
Requirements
- 6+ years of hands-on experience with enterprise-scale networks.
- In-depth knowledge of data centers, servers, and network equipment.
- Experience administering InfiniBand networks using Mellanox/NVIDIA technologies and Ethernet networks using Cumulus/SONiC.
- Strong understanding of L2 and L3 network protocols, network documentation, and diagrams.
- Ability to provide IT support, troubleshoot network errors, and communicate clearly in English.
- Availability for hybrid work in South Korea with 3 days per week in the office and participation in a weekly on-call rotation.
Nice to have
- Experience installing, monitoring, and maintaining large-scale data center networks.
- Experience optimizing low-latency and high-bandwidth networking performance.
- Knowledge of VXLAN, EVPN, BGP, OSPF, NCCL, UCX, MPI, UFM, OpenSM, or NetQ.
Culture & Benefits
- Collaborative environment focused on innovation, ownership, and servant leadership.
- Opportunities for continuous learning and exposure to advanced AI tools and technologies.
- Work with enterprise, sovereign AI, and neocloud infrastructure at global scale.
- After-hours incident response and on-call participation are part of the role.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Nscale
5 часов назад
Senior Network Engineer (AI)
6 дней назад
Network Engineer (Low-Latency Networking)
5 дней назад
Network Reliability Engineer (Go/Python)
16 часов назад
Senior Network Engineer (Python)
110 000 - 160 000$
16 часов назад
Senior Network Engineer
110 000 - 160 000$
1 день назад