Назад
Company hidden
2 дня назад

Senior Site Reliability Engineer (Cloud)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Senior Site Reliability Engineer (GCP/Python): Improving reliability, scalability, and operational excellence for large-scale cloud networking platforms with an accent on automation, observability, and distributed systems. Focus on building operational tooling, optimizing Kubernetes infrastructure, and reducing alert noise in production environments.

Location: Hybrid (Sunnyvale, CA) - Must be on-site 3 days a week. No relocation provided.

Company

hirify.global Inc. (Ruckus Networks) develops large-scale cloud networking platforms that support enterprise customers globally.

What you will do

  • Operate and improve highly available, scalable cloud services and infrastructure on GCP.
  • Troubleshoot production issues across applications, infrastructure, networking, and databases.
  • Build operational tooling and automation using Python to improve efficiency and deployment safety.
  • Enhance observability through metrics, logging, tracing, and alerting using Prometheus, Grafana, and ELK.
  • Define and improve SLIs, SLOs, and operational health metrics.
  • Collaborate with engineering, product, and support teams to ensure operational readiness.

Requirements

  • Must be based in Sunnyvale, CA (Hybrid - 3 days on-site).
  • 5+ years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure.
  • Strong programming skills in Python.
  • Hands-on experience with Google Cloud Platform (GCP), Kubernetes, and containers.
  • Experience troubleshooting distributed systems and Linux systems administration.
  • Proficiency with observability tools like Prometheus, Grafana, OpenTelemetry, or ELK.

Nice to have

  • Experience supporting large-scale SaaS or cloud platforms.
  • Familiarity with Kafka or event-driven architectures.
  • Knowledge of wireless or enterprise networking platforms.
  • Experience improving operational processes and reliability practices.

Culture & Benefits

  • Collaborative engineering environment focusing on innovation and connectivity.
  • Opportunity to work with global teams and diverse perspectives.
  • Culture of ownership, operational excellence, and customer focus.
  • Professional growth opportunities within a global community.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →