9 дней назад
Principal Site Reliability Engineer (AWS)
194 000 - 237 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Site Reliability Engineer (AWS): Designing availability, resiliency, observability, and automation solutions for large-scale financial applications with an accent on scalability, latency, security, and microservice reliability. Focus on building deployment and disaster-recovery tooling, resolving complex production incidents, and leading capacity planning and 24x7 operational support.
Location: Hybrid work in Scottsdale, San Francisco, Chicago, or New York; candidates must be eligible to work in the United States.
Base pay in Phoenix, AZ: $194,000–$237,000 per year, plus eligibility for a discretionary incentive plan and benefits.
Company
provides payment and financial-services technology, including Zelle and Paze, for financial institutions, consumers, and small businesses.
What you will do
- Design and implement software, infrastructure, and resiliency patterns that improve availability, scalability, latency, performance, and security.
- Build automation and tooling for deployments, configuration changes, application management, and disaster recovery.
- Design and promote observability and monitoring systems to detect issues proactively and identify root causes.
- Evaluate application capacity, identify performance bottlenecks, and recommend scaling strategies with Product and Business teams.
- Lead incident response, provide technical guidance and runbooks to Level 1 and Level 2 teams, and collaborate with application development teams on microservice architecture and software lifecycle practices.
- Participate in a 24x7 on-call rotation, develop reusable engineering standards, and coach and mentor team members.
Requirements
- 12+ years of related experience in technical or software development environments, typically with a bachelor’s degree in Business, Computer Science, or a related field.
- Experience leading high-priority incident response, improving root-cause analysis, troubleshooting complex environments, and leading cross-functional teams.
- Hands-on experience with one or more of Python, Go, or Java; Docker; microservices architecture; messaging frameworks such as Kafka, SQS, or JMS; databases such as Oracle, DynamoDB, or Aurora; and caching technologies such as Redis or Memcached.
- Strong understanding of Linux administration and networking fundamentals.
- Experience implementing CI/CD pipelines with Git, Chef, Maven, Jenkins, or similar tools.
- Must independently possess U.S. employment eligibility; visa sponsorship is not available. A background and drug screen are required.
Nice to have
- Programming experience with Java, Ruby, Python, JavaScript, or Go.
- Experience supporting customer-facing production applications in a 24x7 environment.
- Working knowledge of AWS, Docker, Kubernetes, or Swarm.
Culture & Benefits
- Collaborative hybrid office environment.
- Medical, dental, and vision coverage, with HSA or FSA contributions.
- 401(k) plan with a 100% company Safe Harbor Match on the first 6% of deferral.
- Flexible time off for exempt employees, paid holidays, and a paid volunteer day.
- 12 weeks of paid parental leave and family-planning support.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Site Reliability Engineer (AWS)
110 000 - 140 000$
6 дней назад
Sr Staff Site Reliability Engineer (AI)
207 400 - 259 200$
7 дней назад
Site Reliability Engineer (AWS)
180 000 - 220 000$
5 дней назад
Sr. Site Reliability Engineer (Kubernetes/AWS)
8 дней назад
Sr Site Reliability Engineer (AWS)
95 000 - 135 000$
7 дней назад
Site Reliability Engineer - Enterprise Technology
200 000 - 250 000$