обновлено 5 дней назад
Staff/Senior DevOps Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff/Senior DevOps Engineer (Cloud Infrastructure/Distributed Systems): Designing and operating scalable infrastructure, deployment automation, and observability for large-scale microservices with an accent on reliability, high availability, and production operations. Focus on building CI/CD and infrastructure-as-code solutions, leading incident response, and developing capacity planning strategies for distributed services.
Location: Taipei, Taiwan or Tokyo, Japan. The position is ideally based in Taiwan; international candidates are welcome.
Company
is a SaaS company that uses AI to support business decision-making and operates product teams building large-scale distributed systems.
What you will do
- Collaborate with product teams to maintain high availability, reliability, and stability across microservices.
- Design and maintain observability infrastructure for logging, monitoring, and alerting.
- Build and optimize CI/CD pipelines, deployment automation, and infrastructure-as-code solutions.
- Lead incident response, troubleshooting, and post-mortem analysis.
- Drive capacity planning and infrastructure scaling strategies.
- Maintain runbooks and operational documentation while participating in the production on-call rotation.
Requirements
- Bachelor's degree in Computer Science, Computer Engineering, or a related field.
- 5+ years of experience in DevOps, SRE, or infrastructure engineering.
- Strong Linux system administration skills and proficiency in Bash plus Python or Go.
- Experience deploying and operating production services.
- Experience with Git, cloud computing concepts, infrastructure automation, and configuration management.
Nice to have
- Experience with Kubernetes, Docker, Ansible, Terraform, Puppet, or Chef.
- Hands-on experience with Jenkins, GitLab CI, GitHub Actions, or ArgoCD.
- Production experience with AWS, GCP, or Azure.
- Database administration experience with PostgreSQL, MySQL, MongoDB, or Cassandra.
- Experience operating large-scale web services and knowledge of security best practices, including IAM and network security.
Culture & Benefits
- Work as part of product teams focused on large-scale distributed systems.
- Support high-quality services through reliability engineering, automation, and operational excellence.
- Participate in an on-call rotation supporting production systems.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →