5 дней назад
Lead System Engineering (Azure)
158 200 - 237 400$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Lead System Engineering (Azure): Building and operating highly available, scalable cloud and microservices systems on Azure and hybrid infrastructure with an accent on site reliability, monitoring, automation, and incident response. Focus on designing cloud migration architectures, reducing MTTR through observability and automated remediation, and troubleshooting distributed production systems across the stack.
Location: 3400 W. Plano Parkway, Plano, Texas 75075, United States
Salary: $158,200–$237,400 USD annually
Company
AT&T is a large telecommunications and technology company offering enterprise cloud, infrastructure, and digital services.
What you will do
- Review user stories and code changes while supporting site reliability engineering for Azure cloud services.
- Configure monitoring, alerting, dashboards, impact analysis, and automated remediation to improve on-call operations and reduce MTTR.
- Coordinate cloud migration of microservice-based architectures across on-premises and cloud environments.
- Build highly available, scalable, and high-performance cloud infrastructure and provide support for DevOps teams.
- Design and maintain CI/CD automation for microservices using Java, J2EE, Jenkins, Maven, Linux, Docker, and Kubernetes.
- Investigate production incidents, perform root cause analysis, optimize distributed systems, and support the ELK stack and observability platforms.
Requirements
- Bachelor’s degree or foreign equivalent in Applied Computer Science, Electronics Engineering, or Computer Engineering.
- Five years of progressive post-baccalaureate experience in the offered role or a related occupation.
- Experience with Azure, Linux administration, scripting automation, and large-scale distributed systems.
- Experience with ARM, Terraform, Ansible, Git, VSTS, CI/CD pipelines, PowerShell, and Python.
- Experience with RESTful and WebSocket APIs, TCP/IP, routing, load balancing, Docker, Kubernetes, serverless technologies, and Lambda.
- Experience with Log Analytics, Dynatrace, Prometheus, Nagios, Kafka, and production incident troubleshooting.
Culture & Benefits
- Medical, dental, and vision coverage.
- 401(k) plan, tuition reimbursement, and paid time off including at least 23 vacation days and nine company holidays.
- Paid parental leave, caregiver leave, adoption reimbursement, and disability benefits.
- Life insurance, supplemental benefit programs, employee assistance, and wellness programs.
- Employee discounts on eligible AT&T mobility, internet, fiber, and phone services.
- Regular full-time schedule of 40 hours per week.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Site Reliability Engineering Team Lead (Principal SRE, Automotive AI)
132 000 - 211 400$
9 дней назад
Head of Site Reliability Engineering (AI)
195 000 - 285 000$
11 дней назад
Principal Site Reliability Engineer (Kubernetes)
190 000 - 220 000$
10 дней назад
Staff Engineer, Software Engineering (SRE Availability and Incident Management)
100 000 - 230 000$
6 дней назад
Site Reliability Engineer Technical Lead (Kubernetes)
100 000 - 150 000$
11 дней назад
Senior Manager - Site Reliability Engineering (SRE)
9 458 - 16 551$