Назад
Company hidden
1 день назад

Operations Support Engineer - Lead (DevOps)

Формат работы
remote (только USA)/onsite
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Operations Support Engineer - Lead (DevOps): Leading platform operations for NCBI’s GitLab- and Kubernetes-based DevOps platform, supporting developers and production infrastructure with an accent on microservices, reliability, and deployment automation. Focus on debugging distributed systems, maintaining SLOs and error budgets, improving incident response, and building scalable CI/CD and configuration-management solutions.

Location: Bethesda, Maryland, or remote within the United States

Company

hirify.global provides professional services and technology support for the National Center for Biotechnology Information at the National Institutes of Health.

What you will do

  • Lead the platform operations support team for NCBI’s GitLab- and Kubernetes-based DevOps platform.
  • Identify and resolve operational, deployment, and runtime issues in microservice environments.
  • Analyze and debug systems across multiple technologies while keeping applications within error budgets and SLOs.
  • Develop technical solutions, document procedures and SOPs, and compile postmortems with corrective actions.
  • Interview, recommend, train, and support new team members.
  • Provide on-call support to internal developers and staff.

Requirements

  • Bachelor’s degree in STEM or equivalent experience.
  • Strong systems debugging skills and comfort with Linux or the UNIX command line.
  • Experience with programming or scripting and with creating processes, procedures, and SOP documentation.
  • General understanding of TCP/IP, HTTP, and related protocols.
  • Ownership mindset, sound judgment, integrity, responsibility, and strong communication skills.
  • Ability to work effectively with users, peers, management, and teams with varying levels of IT knowledge.

Nice to have

  • Kubernetes, OpenShift, cloud, Linux systems administration, or Site Reliability Engineering experience.
  • GitLab or TeamCity CI, GitOps, ArgoCD, configuration management, and automated CI/CD pipelines.
  • Automation with Bash, Ruby, Python, Go, Java, Scala, Rust, C++, or Perl.
  • Docker, Linux internals, network configuration, distributed systems, and network storage knowledge.
  • Monitoring and alerting tools such as Grafana, Prometheus, OpsGenie, or the TIGK stack.

Culture & Benefits

  • Flexible working hours and remote work options.
  • Technical, scientific, and goal-oriented work supporting biomedical research services.
  • On-site and off-site training, conference attendance, and tuition reimbursement.
  • Medical, dental, and vision coverage, 401(k) with employer contribution, paid holidays, and vacation.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →