5 часов назад
Site Reliability Engineer (TypeScript)
180 000 - 220 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (TypeScript): Improving the reliability, performance, scalability, and security of a mission-critical application across on-premises DoD and AWS environments with an accent on application engineering, observability, and incident response. Focus on designing SLIs and SLOs, building monitoring and alerting with the Grafana stack, automating operational toil, and turning production failures into lasting code fixes.
Location: United States; hybrid in Arlington, VA, with approximately 50% on-site work at customer locations and 50% remote. Candidates outside commuting distance must be willing to relocate; relocation assistance is provided.
Salary: $180,000–$220,000 annually, plus equity.
Company
builds collaboration and AI-powered workflow software for military planning and operational coordination, combining modern software, AI, and real-time collaboration.
What you will do
- Improve application reliability and performance directly in the primarily TypeScript codebase.
- Design and operate monitoring, logging, alerting, and dashboards using Prometheus, Loki, Alloy, and Grafana.
- Define and measure SLIs and SLOs and connect reliability targets to actionable alerting.
- Lead incident response and blameless postmortems, turning root-cause findings into lasting code or process fixes.
- Automate repetitive operational work and support production deployments across on-premises DoD and AWS environments, including air-gapped environments.
- Collaborate with product engineering, platform, security, DevOps, and customer success teams while mentoring engineers.
Requirements
- Active Secret clearance required.
- 5+ years of experience in software engineering, SRE, or a related role, including substantial application development and production delivery.
- Strong TypeScript experience or comparable modern-language experience with willingness to work primarily in TypeScript.
- Understanding of the full software development lifecycle, including design, code review, testing, release, and reliability practices.
- Experience with incident response, root-cause analysis, and implementing durable fixes.
- Ability to collaborate across product, platform, and DevOps teams and communicate context openly.
Nice to have
- Node or modern frontend framework experience, CI/CD pipelines, and testing practices.
- Python, Go, or Bash for tooling and automation.
- Containers, Kubernetes, networking, secure configuration, and AWS or AWS GovCloud.
- Grafana stack, ELK, Datadog, Terraform, Ansible, GitOps, service mesh, or Kubernetes operations experience.
- DoD environments, RMF, STIGs, ICD 503, on-premises virtualization, or relevant AWS/CKA/CKAD certifications.
Culture & Benefits
- Remote-first work environment with flexible hours, while this role requires regular customer-site work.
- Unlimited PTO, annual company retreats, and a $1,000 annual home-office budget.
- Health, dental, vision, and life insurance.
- 401(k) plan with company match and eight weeks of parental leave at 100% regardless of state.
- Blameless postmortems and a culture focused on reliability, shared context, and continuous improvement.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 часов назад
Senior DevOps Engineer (AWS/Kubernetes)
185 000 - 218 000$
5 дней назад
Infrastructure Engineer (AI)
200 000 - 350 000$
7 часов назад
Site Reliability Engineer (AI)
100 000 - 300 000$
4 часа назад
Foundational Software Engineer (Infrastructure)
180 000 - 300 000$
6 часов назад
Senior Site Reliability Engineer (AWS)
140 000 - 180 000$
7 часов назад
Senior Site Reliability Engineer (AWS)
160 000 - 220 000$