Назад
Company hidden
12 дней назад

Sr Staff Engineer (AI/ML)

220 800$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Sr Staff Engineer (AI/ML): Training, optimizing, and deploying large language models and multimodal AI systems for local, on-premises, and edge inference with an accent on fine-tuning, model optimization, retrieval, and production integration. Focus on building reproducible training and inference pipelines, deploying resource-constrained models, and ensuring security, privacy, compliance, and operational performance.

Location: San Jose, California, United States. The role requires satisfactory proof of the right to work in the United States.

Base pay: $220,800 per year. Additional variable or incentive compensation and equity may be available.

Company

hirify.global develops low-power programmable logic, FPGA, CPLD, and programmable power management solutions for customers worldwide.

What you will do

  • Train, fine-tune, and evaluate machine learning models, including LLMs, using SFT, LoRA/QLoRA, and RLHF.
  • Optimize models for local inference through quantization, pruning, distillation, batching, and memory optimization.
  • Deploy models on-premises and at the edge with PyTorch, TensorRT, ONNX, vLLM, or llama.cpp.
  • Build reproducible and scalable training and inference pipelines and integrate locally deployed models into production APIs and internal services.
  • Monitor model performance, drift, latency, and resource utilization in production.
  • Collaborate with software engineers, infrastructure teams, domain experts, and business stakeholders to deliver secure, multimodal, and ensemble AI solutions.

Requirements

  • Master’s or Ph.D. in computer science, engineering, or a related field, or equivalent practical experience.
  • 8+ years of AI and machine learning experience, including at least 3 years with LLMs, code generation, large-scale neural networks, RAG, or AI-powered automation.
  • Hands-on experience with open-weight LLMs and proficiency in Python with PyTorch or TensorFlow.
  • Expertise in vector databases and retrieval models, plus experience deploying models in local, on-premises, or resource-constrained environments.
  • Experience with multi-agent AI systems for autonomous coding tasks, Linux, Docker, and basic cloud or on-premises infrastructure.
  • Must provide satisfactory proof of the right to work in the United States.

Nice to have

  • Experience developing multimodal text-and-image and ensemble models.
  • Experience with LangGraph, CrewAI, AutoGen, or OpenAI Assistants API.
  • Strong research, innovation, mentoring, communication, adaptability, and teamwork skills.

Culture & Benefits

  • Fast-paced, results-oriented, team-first environment.
  • Comprehensive benefits may include healthcare, retirement plans, and paid time off.
  • Potential eligibility for variable or incentive compensation and equity.
  • Background and reference checks are required.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →