Назад
Company hidden
3 дня назад

Site Reliability Engineering Professional (Networking)

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineering Professional (Networking): Supporting and improving BT International’s business-critical network and platform services with an accent on reliability, observability, automation, and incident response. Focus on restoring complex services during major incidents, troubleshooting end-to-end network and cloud environments, and building resilient operational tooling.

Location: Onsite at Adhara Building, Ipswich, United Kingdom

Company

BT International provides secure connectivity and digital services to multinational businesses and organisations as part of hirify.global.

What you will do

  • Support the 24x7 operation of core network and platform services, maintaining availability and performance.
  • Monitor services, identify operational risks, and prevent customer-impacting incidents.
  • Lead technical service restoration during major incidents and resolve complex operational issues.
  • Improve reliability through root cause analysis, problem management, automation, observability, and defect reduction.
  • Collaborate with Engineering, Product, suppliers, and global operational teams on platform resilience and operational readiness.
  • Create and maintain runbooks, service maps, playbooks, handover processes, monitoring, and operational tooling.

Requirements

  • Experience supporting business-critical services in a 24x7 Operations, NOC, Service Operations, or SRE environment.
  • Strong knowledge of IP and Optical networking, WAN technologies, MPLS, BGP, OSPF, Ethernet, SD-WAN, Internet, and cloud connectivity.
  • Experience troubleshooting end-to-end services across network, platform, cloud, and application environments.
  • Knowledge of SRE principles, including reliability, availability, automation, observability, and operational resilience.
  • Hands-on experience with Dynatrace, Splunk, Grafana, ELK, Prometheus, or similar monitoring and observability platforms.
  • Strong incident, problem, change, and major incident management skills, including root cause analysis, communication, and stakeholder management.

Culture & Benefits

  • Supportive 24x7 operational environment focused on customer service excellence and continuous improvement.
  • 10% on-target annual bonus.
  • 24x7 online private GP access for employees and immediate family.
  • Paid carers leave, maternity, paternity, and adoption leave.
  • Pension scheme with 5% employee and 10% employer contributions.
  • Holiday purchase scheme, discounted BT and EE products, and optional healthcare, dental, and gym benefits.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →