5 дней назад
Senior Site Reliability Engineer - Dublin (AI)
85 000 - 115 000€
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer - Dublin (AI) (AWS, Terraform, Python): Scaling and operating the VETRO FiberMap GIS fiber management platform with an accent on reliability, cloud infrastructure, observability, and performance. Focus on automating deployments and remediation, improving distributed-system resilience, optimizing PostgreSQL and spatial-data workloads, and applying AI to troubleshooting and incident analysis.
Location: Dublin, Ireland — hybrid work at the Dublin Center of Excellence
Salary: €85,000–€115,000 per year
Company
develops visual data and network asset management software that helps network operators manage fiber networks from strategic planning through daily operations.
What you will do
- Establish and improve SLIs, SLOs, monitoring, alerting, and overall platform health.
- Build, operate, and optimize highly available cloud infrastructure using infrastructure-as-code and modern deployment practices.
- Improve observability, incident detection and response, root-cause analysis, postmortems, and production on-call procedures.
- Automate deployments, remediation, infrastructure management, capacity planning, and repetitive operational work.
- Scale the FiberMap platform and optimize application, database, infrastructure, and network performance, including large datasets handled by PostgreSQL, NATS, mapping, and tile services.
- Improve CI/CD, developer self-service, security controls, disaster recovery, backups, business continuity, SOC 2 audit evidence, and AI-enabled operational analysis.
Requirements
- 5+ years of experience in SRE, DevOps, platform engineering, cloud infrastructure, or a similar role.
- Experience provisioning, deploying, and supporting highly available cloud infrastructure on AWS or similar platforms.
- Strong knowledge of DevOps, CI/CD, monitoring, logging, observability, and incident management.
- Experience troubleshooting distributed systems across applications, databases, networks, and infrastructure.
- Proficiency with scripting languages such as Python and Bash, plus experience with relational and spatial databases including PostgreSQL and PostGIS.
- Knowledge of data and systems security and compliance frameworks such as SOC 2, with the ability to work across Engineering, Product, Security, and Customer teams.
Nice to have
- Understanding of Terraform and infrastructure-as-code best practices.
- Linux-based system administration experience.
- Familiarity with or interest in Node and Go.
- Strong interest in security best practices, AI, and automation for reliability engineering.
Culture & Benefits
- Hybrid work with a designated productive home workspace for confidential video calls.
- Technology support for setting up a home office and a semi-monthly internet stipend.
- Comprehensive benefits package.
- Professional development opportunities and career growth.
- Collaborative, supportive, inclusive, and learning-oriented work environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Site Reliability Engineer
5 дней назад
Senior Site Reliability Engineer (SRE) - Guidewire Cloud Platform (Application)
Twilio
6 дней назад
Staff Software Engineer (L4)
Anthropic
8 дней назад
Incident Response Manager (AI)
290 000 - 365 000$
11 дней назад
Senior Site Reliability Engineer - Platform Reliability (Resilience)
98 400 - 126 900€