Назад
2 дня назад

Technical AI Lead (EvalOps Platform)

Формат работы
remote (только Vietnam)
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
Vietnam
vacancy_detail.hirify_telegram_tooltipВакансия из Telegram канала -

Мэтч & Сопровод

Покажет вашу совместимость и напишет письмо

Описание вакансии

TL;DR
Technical AI Lead (EvalOps Platform) (Python/RAG/LLM Evaluation): Building a production AI evaluation and observability platform with an accent on LLM evaluation, RAG quality, agentic workflows, and distributed tracing. Focus on leading architecture and engineering delivery, designing regression and hallucination detection systems, and balancing build-versus-integrate decisions for pilot clients.
#vacancy #вакансия

Technical AI Lead (EvalOps Platform) 🚀

If you want to lead the build of a production AI evaluation platform from the first line of code to the first paying clients, this role is for you. You’ll be the founding technical lead, owning architecture, engineering standards, and delivery across a small AI-focused team.

👤 Requirements:

- 5+ years of engineering experience
- 2+ years leading teams of 3–8 engineers through a full product build cycle
- Strong production-grade Python backend: async, FastAPI, APIs, integrations, PostgreSQL, Redis
- Deep understanding of LLM evaluation: LLM-as-judge, calibration, golden datasets, regression testing, hallucination detection
- Production RAG experience: ingestion, chunking, embeddings, retrieval, re-ranking, vector databases
- Experience with AI agents and agentic workflows: ReAct, Plan-and-Execute, supervisor/sub-agent patterns
- LLM observability and distributed tracing experience
- Ability to make and justify architecture and build-vs-integrate decisions
- Strong written technical English
- Comfortable working in an ambiguous, early-stage product environment

⭐ Nice to have:
- OpenAI, Anthropic Claude and open-weight model experience
- MLflow or other experiment/dataset versioning tools
- EU AI Act / NIST AI RMF knowledge
- GDPR / PII handling in AI pipelines
- AI red teaming, jailbreak and prompt-injection testing

📌 Responsibilities:
- Lead architecture and a team of 5–7 engineers
- Build production AI evaluation, RAG quality & monitoring systems
- Own LLM observability, regression testing and cost/latency tracking
- Set engineering standards and work with pilot clients

🌍 Location: Da Nang / Hanoi, Vietnam · Remote eligible

💼 Format: Full-time, permanent

💰 Salary: Discussed based on experience

🚀 Product: New AI evaluation & observability platform, with the first paying pilot targeted by month 3 and commercial beta by month 6.

👉 Apply / contact:

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →

Текст вакансии взят без изменений

Источник -