Site Reliability Engineer (Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Site Reliability Engineer (Kubernetes): Building and operating reliable cloud-native infrastructure for distributed database systems with an accent on automation, observability, scalability, and high availability. Focus on designing resilient Kubernetes and cloud solutions, writing Golang automation, optimizing CI/CD pipelines, and troubleshooting complex production systems.
Location: Remote within the EU timezone, preferably from an EU country
Company
develops a contextual data platform that unifies graph, vector, document, key-value, and search capabilities to support enterprise AI applications.
What you will do
- Design, implement, and maintain infrastructure across AWS and Google Cloud.
- Operate and scale Kubernetes-based distributed database systems.
- Write production-grade Golang code to automate infrastructure management and system operations.
- Improve CI/CD pipelines, deployment processes, monitoring, logging, and alerting.
- Develop disaster recovery, fault-tolerance, and high-availability strategies.
- Troubleshoot infrastructure issues, participate in on-call rotations, and collaborate with development and Customer Success teams.
Requirements
- Professional experience as an SRE or DevOps Engineer in a cloud-native environment.
- Strong Kubernetes experience with large-scale distributed systems.
- Experience with AWS and Google Cloud, networking, security, Linux internals, and Docker.
- Knowledge of CI/CD tools and observability systems such as Jenkins, CircleCI, Prometheus, Grafana, or ELK.
- Programming experience with Golang or Python, plus Git and strong troubleshooting skills.
- Ability to work remotely within the EU timezone, preferably from the EU, independently and collaboratively.
Nice to have
- Experience with distributed databases or large-scale data storage systems.
- Infrastructure-as-Code with Terraform and GitOps experience.
- Strong Golang development experience for automation tools, scripts, or services.
- Cloud security best practices and scripting with Python or Bash.
Culture & Benefits
- Work remotely with an EU-timezone focus.
- Contribute to AI and data infrastructure.
- Collaborate with experienced engineering, marketing, and product professionals.
- Work in an inclusive, growth-oriented environment focused on innovation, customer success, and knowledge sharing.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →