Назад
Company hidden
1 день назад

AI Engineer (Agentic AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK/Singapore/US +1 еще
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Engineer (Agentic AI) (Python/LLM): Building the shared platform layer for extensible, measurable, and self-improving agentic AI across real-time conversational and offline video generation runtimes with an accent on orchestration, retrieval, evaluation, tracing, and guardrails. Focus on designing multi-tenant platform primitives, integrating LLM infrastructure, and solving latency, reliability, and data-isolation challenges across multiple customers.

Company

hirify.global is a global video technology company providing live, on-demand, and real-time video products used by more than 1,000 organizations.

What you will do

  • Build the shared platform layer for agentic AI across real-time conversational and offline video generation runtimes.
  • Develop orchestration capabilities including planning, sub-agents, tool calling, parallel execution, and deadline-aware processing.
  • Build platform primitives such as skills registries, evaluation harnesses, guardrails, integration gateways, memory services, and tracing infrastructure.
  • Develop retrieval infrastructure and RAG capabilities, including indexing, retrieval evaluation, chunking, re-ranking, hybrid search, vector databases, and embeddings.
  • Design high-throughput, low-latency distributed services and APIs for multiple internal teams and external customers.
  • Make build-versus-adopt decisions and integrate open-source AI and platform tooling into production systems.

Requirements

  • 4+ years of experience building production backend or platform systems with distributed services, APIs, asynchronous processing, and multiple consumers.
  • Strong Python experience and hands-on production experience with LLM infrastructure, orchestration, tool calling, retrieval pipelines, context management, prompt chaining, and multi-agent coordination.
  • Experience with LangChain, LangGraph, or equivalent orchestration frameworks.
  • Deep experience with RAG systems, vector databases such as Pinecone, Weaviate, or Qdrant, and embedding models.
  • Strong systems design skills, including registries, gateways, memory services, evaluation harnesses, contracts, and versioning.
  • Understanding of real-time latency constraints, streaming architecture, observability, tenant isolation, and production reliability.

Nice to have

  • Experience evaluating LLM systems with metrics, test harnesses, regression detection, DeepEval, Ragas, or similar tools.
  • Experience with multi-tenant platforms, speech or multimodal AI pipelines, observability, and platform-level tracing.
  • Familiarity with Langfuse, Opik, OpenTelemetry, CrewAI, AutoGen, Semantic Kernel, or MCP.
  • Experience implementing AI guardrails, safety controls, grounding checks, and output validation.

Culture & Benefits

  • Hybrid work model with a flexible working environment.
  • Collaborative culture where initiative and participation are encouraged.
  • Opportunities to work on enterprise-grade AI infrastructure consumed by multiple teams and customers.
  • Global organization with offices in New York, London, Singapore, and Tel Aviv.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →