14 дней назад
Observability Platform Engineer (Datadog)
120 000 - 150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Observability Platform Engineer (Datadog): Building and operating end-to-end observability across cloud and on-premises environments with an accent on Datadog platform engineering, automation, telemetry, and service reliability. Focus on migrating from OpenView, implementing APM, logs, traces, SLOs, RUM, and synthetics, and reducing alert noise while integrating incident workflows with ServiceNow.
Location: New York, NY; hybrid schedule with 2–3 days in the office
Salary: $120,000–$150,000 per year, plus discretionary bonus and benefits
What you will do
- Own, architect, and operate scalable Datadog observability solutions across cloud and on-premises environments.
- Lead the migration from OpenView to Datadog while maintaining monitoring coverage and improving monitoring fidelity.
- Automate monitors, dashboards, alerts, and telemetry configuration using Datadog APIs, Terraform, Python, PowerShell, and Bash.
- Implement APM, distributed tracing, log management, infrastructure monitoring, Network Performance Monitoring, RUM, and Synthetic Monitoring.
- Define SLOs and error budgets, reduce alert noise, and improve incident response and MTTR through correlation, enrichment, and monitor tuning.
- Integrate Datadog with ServiceNow, create runbooks and dashboards, and guide engineering teams on instrumentation and observability standards.
Requirements
- 5+ years of experience in observability, APM, SRE, or platform engineering, including 2–3 years of production Datadog experience.
- Deep knowledge of Datadog APM, Infrastructure Monitoring, Log Management, Synthetics, RUM, SLOs, dashboards, monitors, and alerting.
- Experience with Windows Server and Unix environments, including Linux or Solaris, agent deployment, instrumentation, and OS performance analysis.
- Strong scripting and automation skills with Python, PowerShell, Bash, Datadog API or SDK, and Terraform.
- Experience with AWS, Azure, distributed tracing, metrics pipelines, logging standards, ITSM integrations, and ServiceNow.
- Must have authorization and the right to work in the United States without current or future sponsorship.
Nice to have
- Datadog certifications and experience migrating from OpenView, AppDynamics, or Nagios.
- Familiarity with .NET and C#, CI/CD integrations, synthetic testing, network monitoring, and Datadog NPM or NDM.
- Experience in financial services or other regulated industries.
Culture & Benefits
- Individual contributor role with close collaboration across application, SRE/DevOps, infrastructure, and security teams.
- Opportunity to champion Datadog and OpenTelemetry adoption across the engineering organization.
- Benefits include paid time off, medical, dental and vision insurance, retirement benefits, life insurance, and other eligible-employee benefits.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Staff Platform Infrastructure Engineer (AI)
180 000 - 220 000$
13 дней назад
Platform Engineer (AI)
140 000 - 180 000$
12 дней назад
Site Reliability Engineer - Platform (Azure)
12 дней назад
Senior Platform Engineer (Azure)
142 500 - 190 000$
10 дней назад
Cloud Platform Engineer III (Azure)
101 500 - 131 000$
9 дней назад
Software Engineer, Platform Operations
127 000 - 158 700$