Назад
Company hidden
4 часа назад

Senior Site Reliability Engineer (FedRAMP)

161 900 - 242 900$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (FedRAMP): Designing and operating resilient AWS GovCloud and commercial EKS infrastructure, hybrid edge systems, observability platforms, and compliance automation with an accent on FedRAMP authorization, secure architecture, and production reliability. Focus on leading L3 incident response, defending systems during 3PAO audits, troubleshooting distributed logs, and maintaining NIST SP 800-53 controls across Kubernetes and on-premises environments.

Location: Reston, United States; must be a US citizen residing in the United States

Annual base pay: $161,900–$242,900

Company

hirify.global develops cybersecurity and application delivery technologies that help organizations create, secure, and operate digital applications.

What you will do

  • Provide L3 escalation support for critical production incidents, lead incident bridges, and coordinate resolution across engineering, security, networking, and support teams.
  • Design and operate resilient multi-cluster AWS EKS environments across commercial and GovCloud regions, as well as hybrid networks with on-premises edge servers.
  • Implement FedRAMP controls, prepare audit evidence, support 3PAO audits, and maintain secure audit logging with compliant retention and lifecycle policies.
  • Investigate distributed-system issues using Elasticsearch and Kibana, build operational dashboards, and troubleshoot Vector and Fluentd ingestion pipelines.
  • Develop Terraform infrastructure, GitLab CI/CD and ArgoCD GitOps deployments, and Python or Go automation for provisioning and continuous compliance.
  • Build observability and alerting with Prometheus, Grafana, Alertmanager, and Slack, and facilitate evidence-based post-mortems.

Requirements

  • US citizenship and residence in the United States are required.
  • 5+ years of production experience in Site Reliability Engineering, DevOps, or systems architecture, including high-availability systems and L3 incident escalation.
  • Advanced experience with Elasticsearch, Kibana, KQL, log analysis, dashboards, and distributed-system troubleshooting.
  • Direct FedRAMP, NIST SP 800-53, or SOC 2 Type II audit and compliance experience.
  • Expertise in Kubernetes, AWS EKS/GovCloud, Linux administration, networking, OS hardening, hybrid cloud architecture, and secure multi-tenant systems.
  • Experience with Prometheus, Grafana, Terraform, ArgoCD, GitLab CI, Python or Go, and on-premises edge infrastructure and load balancers.

Culture & Benefits

  • Work within a security-focused environment supporting resilient cloud and edge platforms.
  • Participate in a 24/7 rotational on-call schedule.
  • Receive potential incentive compensation, bonuses, restricted stock units, and employee benefits in addition to base pay.
  • Operate in a diverse workplace focused on improving customer and end-user experiences.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →