3 дня назад
Senior/Lead Site Reliability Engineer (AI/LLM)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior/Lead Site Reliability Engineer (AI/LLM): Operating and designing runtime environments for large-scale online game services with an accent on infrastructure architecture, service availability, operational metrics, and automation. Focus on solving complex reliability problems, researching open-source infrastructure technologies, and applying AI to AIOps, script generation, and root-cause analysis.
Location: Singapore-Guoco Midtown; onsite role
Company
develops and operates online games and digital services, including Eggy Party, Marvel Rivals, UU Accelerator, and Ace Racer.
What you will do
- Manage the operational work of online game services and internal research projects.
- Design and select runtime environments for game servers, including servers, virtualization, cloud services, networks, and databases.
- Establish operational metrics and customized data analysis standards.
- Collaborate with product departments to identify issues, optimize technical architecture, and improve user experience.
- Research open-source software, virtualization, databases, and web services, and develop technical solutions for business implementation.
Requirements
- Bachelor’s degree or above in computer science, networking, communications, automation, or a related field.
- Strong knowledge of Linux, computer network architecture, TCP/IP, and HTTP.
- Proficiency in at least one programming language, such as C/C++, Shell, Python, Golang, Rust, or Java.
- Knowledge of open-source technologies including Linux, Nginx, MySQL, Kubernetes, and Istio.
- Strong logical thinking, communication, learning, teamwork, and problem-solving skills.
- Native or near-native Chinese proficiency is required because some working partners and stakeholders are based in China.
Nice to have
- Experience contributing to open-source communities.
- Hands-on experience applying AI to operations, including AIOps analysis, automated script generation, or root-cause analysis.
- AIOps project experience or demonstrated results using AI tools.
Culture & Benefits
- Work with online game services and internal research projects in a large gaming technology organization.
- Focus on open-source research, technical innovation, and operational automation.
- Collaborate with product departments and stakeholders across game and infrastructure functions.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
7 дней назад
Site Reliability Engineer III
7 дней назад
SRE Monitoring Platform Software Engineer (AI)
8 дней назад
Site Reliability Engineer (Cloud Banking)
4 дня назад
SRE Monitoring Platform Software Engineer (Entry Level)
5 дней назад
Site Reliability Engineer (Azure SaaS)
3 дня назад