6 дней назад
Site Reliability Engineer (SRE/DevOps) - Engineering Productivity
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (SRE/DevOps) - Engineering Productivity (Go/Python/Kubernetes): Building and operating secure, scalable, fault-tolerant infrastructure and internal engineering productivity systems with an accent on automation, observability, reliability, and developer experience. Focus on designing CI/CD platforms, resolving complex infrastructure incidents, and improving production systems across a hybrid cloud environment.
Location: Dublin, Ireland
Company
Develops client-to-cloud networking solutions and software-defined infrastructure for data centers, campuses, and routing environments.
What you will do
- Build, deploy, and operate critical production systems with a focus on scalability, reliability, observability, performance, and security.
- Develop automation that reduces operational toil and improves monitoring, alerting, and automated incident response.
- Create incident response runbooks, triage infrastructure issues, and document postmortems.
- Plan production maintenance windows and coordinate with third-party vendor support.
- Partner with product development teams to identify infrastructure bottlenecks and improve developer workflows.
- Evaluate infrastructure and platform best practices, including open-source system implementations.
Requirements
- BSc or MSc in Computer Science, Engineering, or equivalent work experience, plus at least 3 years of experience.
- Practical knowledge of Go, Python, or shell scripting for implementing automation workflows.
- Linux or UNIX administration and debugging experience.
- Experience operating infrastructure or complex software systems at scale.
- Experience with server provisioning, infrastructure-as-code, and troubleshooting complex software problems.
Nice to have
- Experience managing databases, containerization, virtualization, and storage infrastructure.
- Experience with monitoring stacks such as Prometheus, Grafana, Loki, Tempo, InfluxDB, or Thanos.
- Experience managing Elasticsearch, Artifactory, container registries, CI/CD systems, Perforce, or Gerrit.
- Experience managing large Java applications.
Culture & Benefits
- Work in an engineering productivity team supporting internal product development groups.
- Collaborate with engineers across infrastructure, platform, and software development.
- Contribute to an inclusive environment that values diverse perspectives.
- Join an organization recognized for engineering excellence, diversity, compensation, and work-life balance.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
BI.ZONE
2 часа назад
SRE Engineer (PAM)
1 час назад
Site Reliability Engineer (Cloud/Azure/SRE)
Helsing
6 дней назад
Site Reliability Engineer (AI)
2 дня назад
Site Reliability Engineer (AWS)
18 часов назад
Senior Site Reliability Engineer (SRE)
170 000 - 196 000$
2 дня назад