8 дней назад
Site Reliability SME (AWS)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability SME (AWS): Managing cloud operations and service reliability in a secure AWS GovCloud environment with an accent on incident response, process improvement, and operational reporting. Focus on coordinating technical teams, maintaining SLA-based response, and supporting Windows and Linux systems under dynamic federal mission requirements.
Location: Hybrid role based in McLean, Virginia, United States, with work at the corporate headquarters in Tyson’s Corner, Vienna, VA.
Company
delivers AI, cloud, cybersecurity, and modernization solutions for federal agencies, national security, and health and humanitarian missions.
What you will do
- Provide technical leadership for cloud operations and site reliability in a secure AWS GovCloud environment.
- Facilitate incident management for service disruptions and degradations within established SLAs, including support beyond core business hours during outages.
- Collaborate with operations managers, technical leads, environmental engineers, cybersecurity specialists, and stakeholders to prioritize and complete initiatives.
- Improve and document operational processes, procedures, agile SAFe practices, and team structures.
- Develop metrics and reporting for monthly deliverables.
- Streamline task management and documentation in ServiceNow, Confluence, and Jira.
Requirements
- Hands-on experience as a Cloud Operations Engineer with IT operations experience and technical leadership responsibilities.
- Technical expertise in one or more of Windows, Linux, and AWS.
- Five years of related work experience as a technical leader and a bachelor's degree.
- Must be clearable for DoD Public Trust or higher.
- IAT II certification, such as Security+ or equivalent.
- Strong communication, stakeholder management, customer service, delegation, and requirements-gathering skills.
Nice to have
- AWS SRE experience.
- Strong Windows and/or Linux operating system expertise.
- IAT/IAM III certification.
- AWS certification.
Culture & Benefits
- Medical, dental, and vision coverage.
- 401(k) and IRA retirement plans.
- Life insurance, short- and long-term disability coverage, and family leave.
- Paid vacation, sick leave, and public holidays.
- Training and development opportunities.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
12 дней назад
Senior Staff Site Reliability Engineer
232 338 - 290 422$
9 дней назад
Site Reliability Engineer, Tech Lead (AI)
13 дней назад
AVP SRE and Cloud Solutions (AI)
9 дней назад
Senior DevOps / Site Reliability Engineer (SRE) (Cybersecurity)
165 000 - 215 000$
10 дней назад
Site Reliability Engineer (Healthcare)
175 000 - 200 000$
8 дней назад