Senior Observability Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Senior Observability Engineer (Honeycomb/OpenTelemetry): Building and operating IG’s Honeycomb-centric observability platform and driving OpenTelemetry instrumentation across Java, Python, and C++ microservices with an accent on reliable telemetry pipelines, dataset/query design, and actionable reliability signals. Focus on SLOs, burn rates, alert triggers, and incident response workflows to turn telemetry into faster diagnosis and improved system behavior.
Location: London, United Kingdom
Company
is a leading global fintech company providing online trading platforms and financial education.
What you will do
- Own the Honeycomb-centric observability platform: build, maintain, and evolve it for reliability and scalability in a globally distributed trading environment.
- Define observability standards, data models, and integration patterns for telemetry collection, storage, and querying.
- Drive OpenTelemetry adoption and provide hands-on guidance and reusable instrumentation patterns for services in Java, Python, and C++.
- Improve telemetry coverage with development teams to ensure meaningful traces, metrics, and logs across critical services and user journeys.
- Apply SLO concepts (burn rates and alert triggers) to translate reliability targets into actionable observability signals.
- Participate in incident response rota, use observability tooling to accelerate diagnosis, and lead post-incident reviews to improve both systems and observability coverage.
Requirements
- Hands-on experience with Honeycomb (or similar), including dataset design, query building, and using it for production debugging and reliability analysis.
- Practical experience implementing OpenTelemetry instrumentation in Java, Python, and/or C++, including custom collectors/exporters and sampling strategies.
- Experience in complex distributed microservices environments with high transaction volumes or strict reliability requirements.
- Strong communication and collaboration skills to influence engineering practice without direct authority.
- Practical experience with Terraform for observability infrastructure as code in a cloud environment.
- Ability and willingness to cover UK working hours to support collaboration with London-based engineering teams.
- 5–8 years of relevant experience in observability, SRE, or platform engineering roles.
Culture & Benefits
- Competitive salary and a flexible benefits package on top of salary (12%).
- Private medical cover for you and your family, plus life insurance.
- 25 days holiday, with additional days for birthday and voluntary work (28 total), and the option to buy or sell holiday days.
- Unlimited access to the LinkedIn Learning Platform and onboarding support.
- Enhanced parental pay and leave, plus support for parents.
- Contribution to gym memberships and opportunities to participate in ESG initiatives.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →