Software Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Location: Atlanta, Georgia, United States. Office-first culture with three days per week in the office for most roles.
Annual base salary: $116,475–$174,712.50 USD for candidates based in California, Colorado, Connecticut, Nevada, New York, Rhode Island, and Washington. Additional eligibility may include bonuses, equity, commissions, and benefits.
Company
provides an AI-ready governance platform that combines regulatory intelligence, automation, and connected governance workflows for responsible data and AI use.
What you will do
- Design, deliver, and maintain highly available and performant application platforms with Engineering, Operations, and Product teams.
- Build observability, monitoring, centralized logging, metrics dashboards, and actionable alerting.
- Define and maintain SLIs, SLOs, and error budgets aligned with customer experience.
- Automate operational processes and incident response using Python, Bash, Java, or Ruby.
- Build or integrate AI-assisted incident response systems for root-cause analysis, log summarization, anomaly triage, and alert-quality improvement.
- Review architectures with product teams, identify reliability risks, and share findings with technical leadership.
Requirements
- Bachelor’s degree in computer science, engineering, or a related technical or business field.
- 4+ years of application development experience with Java or an equivalent language, including experience with Spring.
- Experience with AWS, Azure, GCP, or other cloud infrastructure; Kubernetes; container orchestration; distributed systems; and microservices.
- Knowledge of database, network, CPU, JVM, memory, thread, and query performance, with experience using SQL or NoSQL databases.
- Hands-on experience with observability tools such as Datadog, Prometheus, or Grafana, plus CI/CD and infrastructure-as-code tools such as Terraform, Helm, Jenkins, or GitLab.
- Production experience deploying AI systems, including LLM tools, machine learning for anomaly detection, vector databases, embeddings, RAG architectures, and prompt evaluation.
Nice to have
- Experience with chaos engineering tools such as Gremlin or Chaos Monkey.
- Background with high-traffic, product-facing services.
- Experience with incident management platforms such as PagerDuty and Datadog.
Culture & Benefits
- Office-first work environment with regular in-person collaboration.
- Comprehensive healthcare coverage, flexible paid time off, and retirement account support.
- Equity RSUs and annual performance bonus opportunities.
- 14+ weeks of paid parental leave and career development opportunities.
- Company-paid privacy certification exam fees.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →