Назад
Company hidden
6 дней назад

Large Language Model Specialist (LLM Fine-Tuning)

100 000 - 150 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Large Language Model Specialist (LLM Fine-Tuning): Designing and operationalizing fine-tuning workflows for large language models across supervised, preference-based, and reinforcement learning approaches with an accent on dataset construction, evaluation methodology, and production-grade training pipelines. Focus on distributed GPU training, parameter-efficient fine-tuning, safety evaluations, reproducibility, and diagnosing failures in large-scale training jobs.

Location: 100% remote within the United States

Salary: $100,000–$150,000 annually

Company

hirify.global is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

What you will do

  • Design and execute fine-tuning experiments for large language models using supervised learning, DPO, RLHF, and related techniques.
  • Construct, curate, and quality-assure instruction-tuning and preference datasets.
  • Build scalable training pipelines with distributed training frameworks and operate large-scale jobs on GPU clusters.
  • Optimize training stability and throughput using hyperparameter tuning, mixed precision, sequence packing, and efficient attention implementations.
  • Develop automated benchmarks, human evaluations, capability probes, and safety and policy evaluations.
  • Collaborate with product, research, engineering, and platform teams while documenting decisions and mentoring engineers.

Requirements

  • Master’s or PhD in Computer Science, Machine Learning, or a related field, or equivalent experience.
  • Six or more years of combined machine learning research and engineering experience with significant LLM exposure.
  • Strong proficiency in Python and PyTorch, with hands-on experience fine-tuning transformer-based language models at non-trivial scale.
  • Experience with FSDP, ZeRO, pipeline parallelism, RLHF or DPO, evaluation methodology, and GPU cluster operations.
  • Strong written and verbal communication skills and a track record of shipping or publishing impactful LLM work.
  • Applicants must be U.S. citizens, Green Card holders, EAD holders, or H-1B transfer candidates.

Nice to have

  • Publications at top-tier machine learning venues.
  • Experience with multimodal model fine-tuning, synthetic data generation, or dataset distillation.
  • Open-source contributions to LLM training libraries.
  • Experience with responsible AI evaluation and red-teaming.

Culture & Benefits

  • Full-time direct W-2 employment.
  • Remote work within the United States.
  • Collaboration with product, design, engineering, operations, research, and business stakeholders.
  • Opportunities for career growth, technical leadership, code review, design review, and mentorship.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →