7 дней назад
Site Reliability Consultant
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Consultant (Cloud Infrastructure and DevOps): Operating and improving customer infrastructure, automation, monitoring systems, and cloud environments with an accent on reliability, availability, infrastructure architecture, and incident management. Focus on designing cloud infrastructure, investigating complex multi-system failures, automating administrative tasks, and leading client technology road maps.
Location: Canada; remote from home with no daily travel requirement to an office
Company
provides strategic database, analytics, cloud, and digital transformation services for mid-sized and large organizations.
What you will do
- Operate, maintain, and administer customer infrastructure to improve operational efficiency, availability, and visibility.
- Design infrastructure architectures, automation, intelligent monitoring systems, maintenance plans, documentation, and standard procedures.
- Prepare root cause analysis reports for outages and incidents and investigate system and application errors.
- Audit client environments, identify resiliency and automation opportunities, and apply infrastructure best practices.
- Act as a technology leader and advisor for clients, including discussions about technology road maps.
- Participate in an escalation-level on-call rotation and contribute to team documentation and continuous improvement.
Requirements
- Hands-on experience with Google Cloud and AWS, including infrastructure as code with CloudFormation, Terraform, or OpsWorks.
- Strong scripting and administrative automation skills using Python and Scala.
- Solid understanding of microservices and container technologies, with Kubernetes experience required.
- Experience with Linux administration, performance tuning, cloud migration, TCP/IP networking, DNS, NTP, DHCP, SMTP, virtual infrastructure, and hypervisors.
- Experience administering web servers, network load balancers, networking devices, switches, and firewalls.
- Knowledge of DevOps practices, software development lifecycles, Puppet design and deployment, troubleshooting, root cause analysis, scheduling, and task estimation.
Culture & Benefits
- Flexible remote work from home with a stable internet connection.
- Training allowance, professional development days, training opportunities, and certification support.
- Home-office equipment, including a laptop with a choice of operating system, plus an annual workspace budget.
- Annual wellness budget, paid vacation and sick days, and a paid day for volunteering.
- Collaboration with experienced colleagues in a continuous-learning environment.
Hiring process
- Successful applicants must fulfill the requirements necessary to obtain a background check.
- Accommodations are available upon request during the selection process.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →