Назад
Company hidden
10 часов назад

Senior AI Engineer

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Singapore
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior AI Engineer (Python/LLM): Building company-wide AI knowledge infrastructure and context systems with an accent on RAG pipelines, vector databases, scalable backend services, and LLM application architecture. Focus on optimizing retrieval quality, latency, token efficiency, and model inference while delivering production AI tools end to end.

Location: Singapore; on-site

Company

Supporting the company-wide AI transformation through knowledge infrastructure, context systems, and internal AI tools.

What you will do

  • Design and build company-wide AI knowledge infrastructure, including an internal wiki, knowledge base, retrieval layer, and context management system.
  • Develop scalable LLM application architecture with RAG pipelines, vector database integrations, prompt workflows, API services, monitoring, and deployment.
  • Own end-to-end delivery of internal AI tools, including backend architecture, basic frontend integration, testing, deployment, and troubleshooting.
  • Translate ambiguous business needs from business, brand, PR, IR, and leadership stakeholders into practical AI systems and technical roadmaps.
  • Optimize token efficiency, latency, caching, retrieval quality, data architecture, and model inference flows.
  • Evaluate AI coding tools, LLM frameworks, vector databases, and third-party APIs; mentor junior engineers or interns when needed.

Requirements

  • 4–7 years of backend engineering experience, including at least 2 years of hands-on LLM application development.
  • Strong Python backend development skills; knowledge of Node.js or Go is a plus.
  • Strong foundations in algorithms, system design, database design, API architecture, distributed systems, caching, and performance optimization.
  • Production-level LLM application experience covering prompt engineering at scale, model selection, inference pipeline design, or RAG architecture.
  • Hands-on experience with RAG and vector databases such as Pinecone, Weaviate, Chroma, or similar tools.
  • Mandarin fluency and English working proficiency are required. The role requires independent decision-making in ambiguous situations.

Culture & Benefits

  • Full-time, on-site work in Singapore.
  • Hands-on ownership in a fast-moving and ambiguous AI transformation environment.
  • Opportunity to establish technical standards, documentation practices, and reusable engineering workflows.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →