6 дней назад
Site Reliability Engineer - Data, Cloud & Developer Experience (AWS)
140 000 - 225 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer - Data, Cloud & Developer Experience (AWS): Improving the reliability of cloud services and platforms through observability systems, automation, incident response, and SRE practices with an accent on AWS infrastructure, monitoring, alerting, and recovery. Focus on incorporating observability into code and deployment pipelines, implementing instrumentation, scaling SRE systems, and reducing operational impact during incidents.
Location: New York 601 Lex, United States
Expected annual base salary: $140,000–$225,000
Company
is a global alternative asset manager with investment vehicles across private equity, real estate, public debt and equity, infrastructure, life sciences, and other asset classes.
What you will do
- Provide technical leadership in adopting SRE methodologies and evolving reliability standards across the firm.
- Design, implement, and maintain observability systems, instrumentation, monitoring, alerting, and performance insights.
- Incorporate observability standards into code and deployment pipelines.
- Automate the management, maintenance, and scaling of SRE systems with minimal manual intervention.
- Participate in on-call rotations, respond to incidents, and plan and practice disaster recovery.
- Facilitate blameless postmortems and collaborate with service owners, developers, and engineers to improve reliability and operational efficiency.
Requirements
- 5+ years of professional experience in infrastructure, software, DevOps, or platform engineering.
- Automation scripting and troubleshooting skills with languages such as Python, C#, or TypeScript.
- Strong AWS experience; Azure experience is preferred.
- Experience with configuration-as-code, infrastructure management, and CI/CD tools such as Terraform, Puppet, and GitLab CI.
- Hands-on experience with Docker and container schedulers, including AWS ECS and EKS.
- Strong Linux and Windows troubleshooting, networking, observability, incident management, communication, and organizational skills.
Nice to have
- Experience with Azure.
- Experience using coding assistants and chat models such as Anthropic and OpenAI.
Culture & Benefits
- Health benefits including medical, dental, vision, and FSA benefits.
- Paid time off and life insurance.
- 401(k) plan and discretionary bonuses.
- Potential eligibility for equity and other incentive compensation.
- Blameless postmortem culture focused on shared ownership and continuous improvement.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Senior Site Reliability Engineer (Kubernetes)
147 600 - 221 400$
Okta
5 дней назад
Senior TDI Site Reliability Engineer (AWS)
165 000 - 225 600$
7 дней назад
Site Reliability Engineer (AWS)
110 000 - 140 000$
6 дней назад
Site Reliability Engineer
145 000 - 175 000$
7 дней назад
Sr Site Reliability Engineer (AWS)
95 000 - 135 000$
4 дня назад
Sr Staff Site Reliability Engineer (AI)
207 400 - 259 200$