Назад
Company hidden
1 час назад

Senior Machine Learning Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Australia
Вакансия из списка Hirify.GlobalВакансия из Hirify RU Global, списка компаний с восточно-европейскими корнями
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Senior Machine Learning Engineer (AI): Building a high-performance runtime for autonomous AI agents to perceive, reason, and act within the Resolution Platform with an accent on agent orchestration, memory structures, and multi-agent coordination. Focus on designing domain-specialized models via RL on production trajectories and hardening evaluation gates for enterprise-scale deployment.

Location: Hybrid in Melbourne, Australia

Company

hirify.global is a CX company building a Resolution Platform to transition from Systems of Record to Systems of Action using AI-driven automation.

What you will do

  • Advance the iterative planner architecture to handle ambiguous goals and manage memory interference under concurrent sessions.
  • Design and implement multi-agent delegation patterns using the Agent-to-Agent (A2A) protocol.
  • Develop RL training infrastructure, including reward curricula and rollout systems, to create domain-specialized models.
  • Build robust multi-turn evaluation suites and automated trajectory analysis integrated into CI.
  • Implement multi-layered enterprise guardrails and supervisor patterns to prevent tool misuse and hallucinations.

Requirements

  • 5+ years of experience building production ML/AI systems.
  • Hands-on expertise in agent architectures, including planning, tool dispatch, memory, and error recovery.
  • Strong evaluation instincts with a track record of building internal evals to bridge the gap between benchmarks and production.
  • Fluency in Python and PyTorch.
  • Must be based in or able to work from the Melbourne, Australia office (Hybrid).

Nice to have

  • Deep experience with Reinforcement Learning (RL) for language models, specifically reward shaping and online/offline tradeoffs.

Culture & Benefits

  • Hybrid work model combining purposeful in-person collaboration with remote flexibility.
  • Inclusive and fulfilling employee experience within a global organization.
  • Commitment to global diversity, equity, and inclusion.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →