18 дней назад
Site Reliability Engineer (Azure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Azure): Improving the reliability and performance of on-premises and cloud-based production environments with an accent on incident prevention, monitoring, automation, and Azure operations. Focus on leading major-incident resolution, conducting SRE maturity assessments, developing automated remediation, and optimizing system health through SLOs, SLIs, and performance analytics.
Location: 6750 Building, Ayala, Makati City, Philippines; hybrid work model with remote work from home several days a week
Company
is a multinational energy company providing energy products and operational, transactional, processing, consulting, and information technology support across global affiliates.
What you will do
- Prevent incidents by baselining service performance, analyzing operational data, applying lessons learned, and identifying process gaps.
- Lead agile product-line teams in troubleshooting application and critical system performance issues.
- Provide on-call operational support and technical leadership during P1 and major incidents across multiple technologies.
- Assess SRE maturity, identify gaps, and recommend improvement plans based on the SRE framework.
- Improve logging, create trigger-based automated resolutions, and develop scripts that eliminate repetitive operational work.
- Monitor production availability, optimize system performance, and promote DevOps and SRE best practices across stakeholders.
Requirements
- Professional IT experience and a degree in Computer Science, Management Information Systems, or a related field, or equivalent experience.
- Experience with on-premises and cloud architecture, Azure, Windows, Linux, performance monitoring, application and database troubleshooting, API integration, and process automation.
- Knowledge of root-cause analysis, postmortems, ITSM, and incident management.
- Experience with KQL, CI/CD pipelines using Ansible, PowerShell, and shell scripting.
- Knowledge of Agile practices, including SAFe PI Planning, value stream mapping, and scrum ceremonies.
- Strong facilitation, communication, analytical thinking, and technical incident leadership skills.
Nice to have
- Full-stack infrastructure and software engineering experience.
- Knowledge of SLOs, SLIs, SLAs, error budgets, emergency response, demand forecasting, and capacity planning.
- Certifications or training in Azure administration, Azure development, Azure DevOps, networking, security, or Linux/Unix administration.
Culture & Benefits
- Health care coverage for employees and eligible dependents, including medical and optical reimbursements.
- Time off, wellness programs, life insurance, and a long-term savings plan.
- Annual corporate incentive bonus, recognition programs, and career development through training, mentoring, volunteering, and networking.
- Inclusive, collaborative, and high-performing workplace with a focus on diversity and employee development.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →