Назад
Company hidden
14 часов назад

Digital - Principal SRE (AI Engineer)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Digital - Principal SRE (AI Engineer) (AI/SRE): Designing, deploying, and operating AI-driven systems and automation for reliable, scalable digital platforms with an accent on machine learning integration, observability, and cloud-native infrastructure. Focus on building AI provider abstraction layers, defining SLOs and evaluation plans, and solving complex reliability, latency, safety, and cost challenges in production.

Location: Columbus, OH, United States; office workplace type. Applicants must be currently authorized to work full-time in the United States. Visa sponsorship is not available.

Company

hirify.global is a banking organization providing digital platforms and mission-critical financial services.

What you will do

  • Design, develop, and deploy AI-driven systems and automation that improve the reliability and efficiency of digital platforms.
  • Monitor the health, availability, performance, and observability of AI-enabled applications and infrastructure.
  • Integrate machine learning models into production environments and establish SLOs, error budgets, and incident response procedures.
  • Investigate complex AI-system incidents, conduct post-incident analysis, automate manual operations, and optimize performance.
  • Develop abstraction layers across AI providers such as Google and OpenAI to support agent workflows.
  • Run design workshops and proofs of concept, and define metrics, test harnesses, and evaluation plans for agent accuracy, latency, safety, and cost.

Requirements

  • Bachelor’s degree in computer science, engineering, data science, or a related field, plus relevant experience.
  • 5+ years of hands-on experience in AI/ML engineering, SRE, DevOps, or related roles.
  • Programming experience in Python, Java, or similar languages, including developing and deploying machine learning models.
  • Experience with AWS or GCP, Docker, Kubernetes, infrastructure as code using Terraform or Ansible, and CI/CD pipelines.
  • Familiarity with Prometheus, Grafana, the ELK stack, and ServiceNow incident management.
  • Full-time United States work authorization is required; immigration sponsorship is not available.

Nice to have

  • Experience operationalizing large language models or generative AI systems in production.
  • Background in MLOps, data engineering, or cloud-native AI deployment.
  • Knowledge of security best practices for AI and cloud infrastructure.
  • Contributions to open-source AI or SRE projects and technical communities.

Culture & Benefits

  • Collaboration with Digital SRE engineers, data scientists, DevOps, operations teams, and cross-segment stakeholders.
  • Opportunities to work on AI solutions and reliability for mission-critical digital services.
  • Focus on reusable patterns, documentation, runbooks, and knowledge sharing.
  • On-call responsibilities may be required.
  • Exempt position and not eligible for overtime pay.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →