4 часа назад
Senior Site Reliability Engineer (AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AWS/Kubernetes): Modernizing and operating a cloud platform across AWS, Kubernetes, CI/CD, observability, infrastructure as code, and engineering tooling with an accent on reliability, security, and infrastructure efficiency. Focus on optimizing distributed systems at scale, automating service-team workflows, building AI-enabled SRE tooling, and mentoring engineers.
Location: Tokyo, Japan; hybrid work with 3 days per week in the Tokyo office. Up to 5% travel may be required.
Company
An AI product company building an agentic experience platform that helps marketing and data teams automate ROI-driven customer engagement across web, mobile, and desktop.
What you will do
- Help define and execute the technical roadmap for the AWS cloud platform, including EC2, Kubernetes, CI/CD, observability, infrastructure as code, and engineering tooling.
- Collaborate with service teams to automate workflows, drive adoption of platform releases, and contribute improvements across codebases.
- Optimize AWS networking, storage, compute, and orchestration layers to improve efficiency and reduce costs.
- Partner with security and IT teams to identify vulnerabilities, automate access management, and maintain strong security standards.
- Explore AI applications in SRE and engineering, building tooling and workflows for organizational adoption.
- Advise new projects on reliable, efficient architecture and mentor junior engineers.
Requirements
- 5+ years of experience in SRE, DevOps, software engineering, or a related role working with distributed systems at scale.
- Hands-on experience with AWS and related cloud technologies such as EKS, EC2, and infrastructure as code.
- Proficiency in at least one programming language, such as Python or Go.
- Experience with modern SaaS development practices, including Git and CI/CD.
- Strong English communication skills are required for collaboration in a diverse, distributed environment.
Nice to have
- Advanced knowledge of AWS cloud networking.
- Understanding of Agile practices such as Scrum or Kanban.
- Open-source community contributions.
- Conversational Japanese skills.
Culture & Benefits
- Work with an approximately 10-person Japan SRE team within Engineering Services.
- Collaborate asynchronously with a North American SRE team through Slack, Jira, and Confluence.
- Share a backlog and on-call rotation with the North American team to provide around-the-clock coverage.
- Work in an environment focused on diversity, equity, inclusion, belonging, ownership, and continuous innovation.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
DeepL
6 дней назад
Senior Platform Engineer (Kubernetes)
2 дня назад
SRE (Infrastructure) Lead Engineer
6 дней назад
IT Staff Systems Engineer - VM (Infrastructure Automation & Virtualization)
DeepL
5 дней назад
Developer Experience Engineer (AI)
18 часов назад
Senior Observability & Telemetry Engineer (GPU/AI Infrastructure)
5 дней назад