Назад
Company hidden
7 часов назад

Staff II Software Engineer AI/ML Ops (AI)

245 000 - 307 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff II Software Engineer AI/ML Ops (AI/ML): Building and operating production ML and LLM agent systems, model registries, feature stores, experimentation frameworks, and scalable cloud infrastructure with an accent on reliability, governance, observability, and secure deployment. Focus on architecting agent guardrails, automating model and agent CI/CD, implementing evaluation and monitoring systems, and optimizing distributed training and inference workloads.

Location: Pleasanton, United States; candidates within a reasonable commute to a hirify.global office work in the office at least 3 days a week.

Salary: $245,000–$307,000 per year, plus eligible short-term and long-term incentive programs and benefits.

Company

hirify.global is a technology company providing finance automation products and developing AI-driven accounting agents.

What you will do

  • Lead the development and optimization of PySpark ETL pipelines and integrations with client data sources, APIs, Fivetran, Plaid, and internal connectors.
  • Architect and operate ML and LLM production infrastructure, including model registries, feature stores, distributed training, GPU/TPU allocation, and cloud scaling.
  • Build CI/CD pipelines for model and LLM agent deployment, validation, prompt evaluation, rollback automation, and reproducible experimentation.
  • Define agent runtime governance, MCP Registry lifecycle management, authentication, authorization, session governance, and secure tool usage.
  • Implement observability, logging, metering, auditing, evaluation, and feedback systems for model and agent quality, safety, latency, and compliance.
  • Partner with data science, security, product, and engineering teams; lead design reviews, incident response, reliability strategies, and mentoring.

Requirements

  • Extensive practical experience delivering complex business solutions and operating with a high degree of autonomy as an expert individual contributor.
  • Bachelor’s or Master’s degree in Computer Science, Machine Learning, Data Science, or a related field.
  • Strong programming skills in Python, Java, or Scala, with experience in ML frameworks such as TensorFlow, PyTorch, or scikit-learn.
  • Experience operating production ML and LLM pipelines across GCP, AWS, or Azure, including Airflow, Kubeflow, MLflow, Vertex AI, CI/CD, IaC, and DevSecOps.
  • Experience with containers and infrastructure, including Docker, Kubernetes, cloud optimization, networking, security protocols, and automation scripting.
  • Familiarity with LangChain, LangGraph, ADK or similar agentic runtimes, plus Prometheus, Grafana, New Relic, or comparable observability stacks.

Culture & Benefits

  • Technology-focused environment centered on finance automation and future-oriented product development.
  • Kind, open, and accepting culture that values diverse perspectives.
  • Professional development seminars and learning opportunities.
  • Inclusive affinity groups supporting diversity and employee growth.
  • Benefits and wellness plans, with eligible short-term and long-term incentive programs.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →