9 дней назад
Staff Software Engineer, Platform Infrastructure (Go)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Software Engineer, Platform Infrastructure (Go): Building an internal Service Foundation platform and reusable infrastructure patterns for reliable, observable services with an accent on SRE practices, GCP managed services, Terraform, and security compliance. Focus on establishing SLO/SLI frameworks, designing Prometheus/OpenTelemetry observability, and scaling paved paths across 15+ engineering teams.
Location: Remote within Ireland. Scheduling flexibility is expected for cross-timezone standups and incident response with teams in North America and Europe.
Company
operates a WebOps platform used to build, iterate, and scale WordPress, Drupal, and Next.js websites for organizations worldwide.
What you will do
- Establish SRE practices across the Platform Infrastructure Engineering team and the broader Internal Platform Group, including SLOs, SLIs, error budgets, incident response, and toil reduction.
- Define observability standards using Prometheus metrics, OpenTelemetry tracing, structured logging, and Grafana.
- Build reusable service templates, Terraform modules, Go APIs, CLIs, Cloud Run services, and GKE workloads with reliability and observability included by default.
- Lead adoption of GCP-native managed services, including Secret Manager, Cloud SQL, Cloud Run, GKE, GitHub Actions, and Cloud Build.
- Apply security and compliance practices across IAM, IAP, VPC design, networking, secrets management, SOC 2, ISO 27001, and PCI DSS requirements.
- Mentor engineers, provide technical direction, operate systems through their full lifecycle, and join the on-call rotation after an initial 3–6 month ramp period.
Requirements
- 8+ years of experience building and operating production systems, with significant infrastructure, SRE, or platform engineering experience.
- Deep knowledge of SRE principles, including SLOs, SLIs, error budgets, incident management, on-call practices, and measurable reliability improvements.
- Strong hands-on experience with Go 1.21+, Python, GCP, Cloud Run, GKE or equivalent Kubernetes platforms, Cloud SQL, IAM, networking, and managed services.
- Strong Terraform skills, including designing reusable modules and infrastructure patterns adopted across teams.
- Production experience with Grafana, Prometheus, OpenTelemetry, metrics, traces, logs, IAP, VPC design, and security or compliance frameworks.
- Experience setting technical direction, building self-service developer platforms, influencing engineers without direct authority, and communicating reliability risks clearly.
Culture & Benefits
- Remote-forward, distributed engineering environment with a focus on collaboration, autonomy, character, and a no-blame culture.
- 28 days of holiday, private medical and dental coverage, life and critical illness insurance, and a workplace pension scheme.
- Industry-competitive compensation with an equity plan.
- Top-of-line equipment, wellness and reading allowance, and access to LinkedIn Learning.
- Team and company-wide events and activities.
- Visa sponsorship is not available.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
Senior Platform & Reliability Engineer (Kubernetes)
11 дней назад
Senior Software Engineer, Platform Infrastructure (Remote)
158 000 - 190 000$
12 дней назад
Staff Software Engineer, Infrastructure (Blockchain)
14 дней назад
Principal Platform Engineer (AI/Identity Platform)
10 дней назад
Senior Platform Engineer
IBM Watson
11 дней назад
Staff Software Engineer - Infrastructure (Kubernetes)
161 000 - 299 000$