Назад
Company hidden
1 день назад

Forward Deployment Engineer (AI)

138 000 - 170 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Forward Deployment Engineer (AI): Designing, building, and deploying production generative AI applications on SambaNova's SN40L platform and SambaStack portfolio with an accent on LLM orchestration, RAG pipelines, fine-tuning workflows, and inference optimization. Focus on benchmarking throughput, latency, and accuracy, troubleshooting across model, software, and hardware layers, and translating customer requirements into scalable AI deployments.

Location: Remote - US; willingness to travel up to 50% to customer sites.

Salary: $138,000–$170,000 USD base salary per year.

Company

hirify.global develops an integrated hardware and software platform for enterprise and government generative AI and high-performance computing.

What you will do

  • Embed with strategic enterprise customers to design, build, and deploy production GenAI applications on the SN40L platform and SambaStack.
  • Architect LLM-powered workflows, including RAG pipelines, multi-agent systems, fine-tuning workflows, and coding solutions.
  • Optimize inference performance on hirify.global hardware and benchmark model throughput, latency, and accuracy.
  • Troubleshoot production issues across model, software, and hardware layers as the technical escalation lead.
  • Translate customer needs into product requirements and engineering feedback while partnering with Product and Engineering.
  • Develop reusable accelerators, reference architectures, and playbooks, and present technical findings to customers and internal stakeholders.

Requirements

  • 5+ years of hands-on engineering experience shipping production AI/ML systems.
  • Deep expertise in LLM orchestration, RAG, agentic frameworks, prompt engineering, and evaluation pipelines.
  • Strong knowledge of model training, fine-tuning, inference optimization, quantization, and performance benchmarking.
  • Proficiency in Python; working knowledge of C++ or CUDA is advantageous for hardware-layer debugging.
  • Experience with AWS, Azure, or GCP, plus containerization, Kubernetes, Docker, and MLOps tooling.
  • Bachelor's or graduate degree in Computer Science, Electrical Engineering, Mathematics, Physics, or equivalent practical experience; strong customer-facing communication skills.

Nice to have

  • Experience with AI accelerators or custom silicon, including TPUs.
  • CUDA or low-level GPU programming experience.
  • Familiarity with vLLM or SGLang.
  • Enterprise AI deployments in regulated industries.

Culture & Benefits

  • Full-time US employment with equity and a competitive total rewards package.
  • Medical insurance with 95% employee premium coverage and 77% dependent premium coverage.
  • Dental, vision, disability, life, AD&D, HSA, and FSA benefits.
  • Well-being benefits including Headspace, Gympass+, One Medical, and counseling through an Employee Assistance Program.
  • Opportunity to work on enterprise-grade generative AI infrastructure and represent the platform at industry events.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →