3 дня назад
Platform Engineer (Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Platform Engineer (Kubernetes) (Infrastructure/SRE): Running and evolving observability, GitLab, CI runner, and platform services with an accent on reliability, automation, and self-service. Focus on managing production incidents, delivering infrastructure as code, performing Kubernetes upgrades through GitOps, and using AI engineering assistants safely in production workflows.
Location: Fully remote; work from any location worldwide
Company
builds Linux infrastructure and security products, including platform, automation, and management services.
What you will do
- Run the observability platform, onboard teams, and manage alerting, cost, and capacity.
- Operate GitLab and the CI runner fleet, including upgrades, access, backups, and restore drills.
- Keep production services reliable through monitoring, runbooks, repeatable recovery, and infrastructure as code.
- Research, design, and deploy new platform services with appropriate monitoring, backups, documentation, and automation.
- Respond to developer requests, convert recurring work into self-service, and lead incident response, root-cause analysis, and post-mortems.
- Use AI engineering assistants for research, planning, and automation while reviewing and validating their output before production use.
Requirements
- Senior-level infrastructure, platform, or site reliability engineering experience with ownership of production services.
- Linux administration and debugging on bare metal and virtual machines, plus production Kubernetes delivered through GitOps.
- Experience with Kubernetes cluster upgrades, Ansible, Terraform or OpenTofu, and merge-request-based infrastructure delivery.
- Production GitLab administration and GitLab CI experience, plus working knowledge of Prometheus, Grafana, alert rules, dashboards, and PromQL.
- Strong written communication and interpersonal skills for working with product and engineering teams.
- English at upper-intermediate level or higher, with advanced experience using AI engineering assistants such as Claude and Codex.
Nice to have
- SLOs, burn-rate alerts, and data-driven alerting design.
- Kata Containers, Firecracker, gVisor, or other microVM isolation technologies.
- Ceph RGW or similar S3-compatible object storage operations.
- AWS cost management, self-hosted Sentry, Kafka, ClickHouse, Redis, Python, or Go experience.
Culture & Benefits
- Fully remote work with flexible working hours.
- Professional development opportunities and an education budget.
- 24 paid vacation days, 10 national holidays, and unlimited sick leave.
- Private medical insurance compensation plus coworking and gym or sports reimbursement.
- Opportunity to receive a reward for an innovative idea that the company can patent.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 дней назад
Platform Engineer (Kubernetes)
6 дней назад
Platform Engineer, DevOps (AI)
6 дней назад
AI Platform Engineer (AI)
5 дней назад
OpenShift Platform Engineer (Kubernetes)
86 000 - 103 000$
9 дней назад
Software Engineer, Platform Operations
127 000 - 158 700$
Kaufland e-commerce
8 дней назад