обновлено 5 часов назад
Site Reliability Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Cloud/Kubernetes): Building and operating KNIME’s next-generation managed SaaS cloud platform with an accent on reliability, security, scalability, and cost efficiency. Focus on infrastructure as code, production readiness, active on-call ownership, incident mitigation, observability, and automated operations for a multi-tenant platform.
Location: Berlin or Konstanz, Germany; hybrid working
Company
develops a data science and analytics platform and is building its next-generation managed cloud products and services.
What you will do
- Automate the deployment and operation of large-scale SaaS systems using code.
- Build infrastructure as code with Helm, Terraform, Amazon CloudFormation, and Azure ARM.
- Participate in on-call rotations, incident triage, live-environment troubleshooting, root cause analysis, and issue resolution.
- Set and promote standards for reliability, scalability, traceability, monitoring, and production readiness across product and engineering teams.
- Instrument deployed systems for performance, reliability, and cost effectiveness.
- Lead planning, manage dependency risks, and drive consistent adoption of operational standards.
Requirements
- Current certification in a cloud platform or related technology such as AWS, Kubernetes, or Linux.
- Strong experience with AWS or Azure; knowledge of services including VPC, IAM, EKS, ECR, EC2, S3, RDS, and CloudWatch, or their Azure equivalents.
- Experience deploying software systems to Kubernetes and applying common Kubernetes deployment patterns.
- Required scripting experience with Python and Shell; additional Go or Java experience is beneficial.
- Systems-level Linux knowledge, including networking, security, routing, load balancers, and firewalls.
- Knowledge of service telemetry, distributed logging, tracing, OAuth/OIDC identity providers such as Keycloak, and relational databases such as Postgres.
Nice to have
- Experience with Kubernetes operators.
- Programming experience in Go or Java.
- Experience working independently and communicating across geographically and culturally dispersed teams.
Culture & Benefits
- Purpose-driven work shaping the next generation of cloud service offerings.
- Systems-first culture focused on simplicity, maintainability, and clean design.
- Hands-on learning, peer exchange, and exposure to AI and data analytics topics.
- Hybrid working and flexible hours.
- Subsidised sports or yoga courses, physiotherapy, and flu shots at select locations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →