4 дня назад
Senior Site Reliability Engineer (SRE) – Infrastructure & Systems
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (SRE) – Infrastructure & Systems (Kubernetes): Owning and evolving Webshare’s production infrastructure across hundreds of servers and dozens of services with an accent on Kubernetes migration, high availability, and observability. Focus on leading the Docker Swarm-to-Kubernetes migration, strengthening IaC and CI/CD reliability, and resolving complex infrastructure incidents.
Location: Tallinn, with hybrid work (3 office days and 2 work-from-home days) and work-from-anywhere options.
Company
Webshare is an enterprise-grade proxy platform providing self-service access to more than 80 million global IPs across 195 countries through developer-friendly APIs.
What you will do
- Own and evolve production infrastructure across hundreds of servers and approximately 50 services.
- Lead the migration from Docker Swarm to Kubernetes in on-premises and bare-metal environments.
- Maintain high availability and improve observability in collaboration with backend engineers.
- Establish Infrastructure as Code practices and improve CI/CD pipeline reliability.
- Participate in on-call rotations, lead infrastructure incident resolution, run post-mortems, and drive systematic remediation.
- Build platform tooling that improves developer experience and reduces infrastructure toil.
Requirements
- Experience building and operating highly available infrastructure at production scale, including hundreds of servers and dozens of services.
- Hands-on Kubernetes experience in self-hosted or bare-metal environments.
- Strong Infrastructure as Code, scripting, and development skills.
- Experience leading at least one major infrastructure migration from planning through stabilization.
- Familiarity with Python and/or Go, proxy and networking-heavy infrastructure, and edge or split compute architectures.
- Experience working in a small team where developers share infrastructure responsibility.
Nice to have
- Backend experience with Python or edge-service experience with Go.
- Exposure to proxy infrastructure, edge clusters, or split compute/edge architectures.
Culture & Benefits
- Hybrid work with 3 office days and 2 work-from-home days, plus work-from-anywhere options.
- Internal and external learning resources, guilds, communities, and knowledge sharing across the Tesonet network.
- Yearly workation, internal and community events, and a team-building budget.
- Private health insurance, extra days off, and family-related time off.
- Access to office perks, the latest tools and gadgets, a 24/7 gym, physical well-being specialists, and psychologist or psychotherapist sessions.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
Staff Site Reliability Engineer (AI/ML)
9 дней назад
Staff Observability Engineer (AI)
290 000 - 375 000PLN
6 дней назад
IT Operations Engineer
800 - 950PLN
Bolt
9 дней назад
Senior Site Reliability Engineer (Database Reliability)
10 дней назад
Manager, Site Reliability Engineering (Cloud Infrastructure)
10 дней назад