2 дня назад
Site Reliability Engineer (DataCosmos)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (DataCosmos) (Cloud/Kubernetes): Ensuring the reliability, scalability, and performance of a geospatial data platform and its processing pipelines with an accent on observability, high availability, and safe deployments. Focus on monitoring distributed systems, responding to incidents, improving CI/CD, and automating operational processes.
Location: Hybrid in Harwell, UK; Barcelona, Spain; Athens, Greece; or Porto, Portugal. The role can be based in any European location, and applicants must have the legal right to work in the chosen location.
Company
develops satellite missions, Earth observation, IoT connectivity, and secure communications solutions using actionable data from space.
What you will do
- Own the reliability, performance, and scalability of the DataCosmos data platform and processing pipelines.
- Monitor infrastructure and data flows end to end to maintain system visibility.
- Respond to incidents, troubleshoot operational issues, and implement long-term fixes.
- Improve deployment processes and contribute to CI/CD pipelines for safe, repeatable releases.
- Design resilient, scalable systems with engineering teams and automate operational processes.
- Support customer-impacting issues in collaboration with Customer Success.
Requirements
- Strong experience working with Linux systems and cloud platforms such as AWS, GCP, or Azure.
- Solid Kubernetes knowledge and experience running production systems.
- Understanding of observability, including monitoring, logging, and tracing.
- Ability to design or operate highly available, distributed systems.
- Focus on automation, scalability, and continuous improvement.
- Legal right to work in the selected European location. Submit the CV in English.
Culture & Benefits
- Work on space technology and satellite data products serving customers globally.
- Contribute to a mission focused on making space more accessible and solving real-world challenges.
- Join a diverse, ambitious, and supportive team.
- Work in a fast-moving environment where reliability is critical.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
Lead Site Reliability Engineer
Datadog
3 дня назад
Staff Engineer, Compute (Kubernetes)
9 часов назад
Senior Site Reliability Engineer (Kubernetes)
2 дня назад
Site Reliability Engineer (Cloud/SaaS)
Datadog
3 дня назад
Staff Engineer, Compute (Kubernetes)
2 дня назад