9 часов назад
Sr. Site Reliability Engineer
130 000 - 140 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Sr. Site Reliability Engineer (AWS/Azure/Kubernetes): Maintaining and improving reliable, scalable production systems for financial services and healthcare technology with an accent on incident response, observability, automation, and infrastructure resilience. Focus on building diagnostic and monitoring tools, validating production changes, analyzing root causes, and ensuring availability through on-call operations and failover testing.
Location: Waltham, Massachusetts, United States — hybrid at 10 CityPoint
Salary: $130,000–$140,000 per year
Company
SS&C provides proprietary technology and financial services and healthcare solutions for organizations in 35 countries.
What you will do
- Investigate customer, database, application support, and monitoring incidents and develop short- and long-term solutions.
- Support production information systems, deployments, releases, and change-request implementations.
- Respond to escalated incidents within SLA and participate in production issue bridges with R&D, architecture, and operations teams.
- Build diagnostic, analytical, health-monitoring, and API tools to improve MTTA, MTTD, MTTR, availability, scalability, and latency.
- Validate software deliverables for production readiness and assess and mitigate production change risks.
- Participate in on-call coverage and test network redundancy, resilience, and failover.
Requirements
- BS or MS in Computer Science or a related discipline.
- Strong experience with Unix/Linux, Java web-based enterprise applications, Python, or Bash.
- Strong troubleshooting experience with Kubernetes and Docker.
- Experience with AWS, Azure, AWS services including CloudWatch, EKS, EFS, S3, and Redshift, and IaC tools such as Terraform or Ansible.
- Knowledge of networking and application and transport protocols, including HTTP(S), JMS, TCP, and UDP.
- Experience with monitoring tools, DNS/CDN/WAF platforms, databases, and messaging systems such as Splunk, Datadog, Dynatrace, Prometheus, Akamai, Cloudflare, Oracle, PostgreSQL, MongoDB, RabbitMQ, Interconnect, or AMQ.
Nice to have
- At least 7 years of experience developing software projects and/or working in DevOps or SRE.
- Familiarity with AI tools.
Culture & Benefits
- Work with financial services and healthcare technology platforms in large-scale enterprise environments.
- Medical, dental, and vision coverage.
- 401(k) plan with company match.
- Paid time off, holidays, and parental leave.
- Professional development reimbursement.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
1 день назад
Site Reliability Engineer
105 600 - 145 200$
2 дня назад
Sr. DevSecOps Engineer II
170 000 - 190 000$
5 дней назад
Sr. Site Reliability Engineer, Data - FreeWheel
117 628 - 176 442$
2 дня назад
Sr. Staff Site Reliability Engineer (Federal, AI)
140 000 - 200 000$
9 часов назад
Staff Site Reliability Engineer (AWS GovCloud)
158 500 - 230 000$
6 дней назад
Site Reliability Engineer
87 400 - 123 400$