4 дня назад
Senior Site Reliability Engineer (AWS/AI)
140 000 - 160 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AWS/AI): Scaling and operating the VETRO FiberMap GIS platform with an accent on cloud infrastructure, reliability engineering, observability, and performance optimization. Focus on automating deployments and remediation, improving distributed-system resilience, optimizing large spatial datasets, and using AI to reduce operational toil.
Location: Portland, Maine, United States — hybrid work at the Portland HQ with remote and in-person work
Salary: $140,000–$160,000 USD annually
Company
develops FiberMap, a GIS fiber network management platform that gives network operators visibility and control over physical network assets.
What you will do
- Establish and improve SLIs, SLOs, monitoring, alerting, observability, and platform health.
- Build, operate, and optimize highly available cloud infrastructure using infrastructure as code and modern deployment practices.
- Improve incident detection, response, root-cause analysis, postmortems, on-call procedures, and prevention of recurring incidents.
- Automate deployments, remediation, infrastructure management, capacity planning, and repetitive operational work.
- Scale and optimize applications, databases, networks, and large datasets using PostgreSQL, NATS, mapping, and tile services.
- Improve CI/CD, developer self-service, security, disaster recovery, backups, business continuity, SOC 2 readiness, and AI-enabled operations.
Requirements
- 5+ years of experience in SRE, DevOps, platform engineering, cloud infrastructure, or a similar role.
- Experience provisioning, deploying, and supporting highly available cloud infrastructure on AWS or similar platforms.
- Strong knowledge of DevOps, CI/CD, monitoring, logging, observability, incident management, and distributed-system troubleshooting.
- Proficiency with scripting languages such as Python and Bash, plus experience with relational and spatial databases including PostgreSQL and PostGIS.
- Understanding of data and systems security and compliance frameworks such as SOC 2.
- Ability to work collaboratively across Engineering, Product, Security, and Customer teams in a growth-stage environment.
Nice to have
- Experience with Terraform and infrastructure-as-code best practices.
- Linux system administration experience.
- Familiarity with or interest in Node and Go.
- Interest in applying AI and automation to reliability engineering.
Culture & Benefits
- Hybrid work with a designated productive home-office space for confidential video calls.
- Home-office technology assistance and a semi-monthly internet stipend.
- Health insurance and retirement plans with a 401(k) match.
- Remote work stipend, Learning & Development fund, flexible vacation, and PTO.
- Professional development and career growth opportunities in a rapidly growing company.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Site Reliability Engineer (AI Infrastructure)
175 000 - 265 000$
9 дней назад
Staff Site Reliability Engineer (Cybersecurity)
199 750 - 270 000$
5 дней назад
Senior Site Reliability Engineer I (AI)
95 300 - 158 800$
10 дней назад
Principal Site Reliability Engineer (Kubernetes)
190 000 - 220 000$
Baseten
7 дней назад
Site Reliability Engineer (AI)
165 000 - 330 000$
Okta
5 дней назад
Staff Site Reliability Engineer, Federal (TS/SCI)
174 000 - 238 000$