Назад
Company hidden
3 дня назад

Lead Software Platform Engineer (AI/ML)

200 000 - 270 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Lead Software Platform Engineer (AI/ML): Building and scaling multi-tenant AI/ML infrastructure for scientific data platforms with an accent on distributed systems, MLOps, model serving, and production LLM workflows. Focus on designing secure inference, evaluation, observability, and agent platforms that meet scalability, reliability, cost, and regulatory requirements in pharmaceutical environments.

Location: Remote in the United States; company locations include Cambridge, Massachusetts and San Mateo, California.

Salary: $200,000–$270,000 USD per year

Company

hirify.global develops scientific data and AI cloud infrastructure and laboratory data management solutions for pharmaceutical and scientific organizations.

What you will do

  • Own the architecture and API surface of a multi-tenant AI/ML platform used by customers and internal engineering teams.
  • Manage model and prompt lifecycles across Databricks MLflow and AWS Bedrock, including versioning, promotion, rollback, and multi-model serving.
  • Design real-time and batch inference infrastructure with routing, batching, caching, concurrency controls, accelerator capacity planning, and graceful degradation.
  • Build production LLM and agent capabilities using RAG, tool and function calling, MCP-based tooling, and agent runtimes.
  • Establish security, tenant isolation, evaluation, observability, reproducibility, lineage, and auditability for AI systems in regulated pharmaceutical environments.
  • Set technical direction, lead design reviews, support production readiness and incident response, and mentor engineers across distributed systems and AI engineering.

Requirements

  • 10+ years of software and infrastructure engineering experience building and scaling distributed cloud-native systems in production.
  • Technical leadership or architecture experience with accountability for system design, scalability, performance, and cost optimization.
  • Production experience building multi-tenant AI/ML infrastructure for external users, including model deployment, lifecycle management, and serving.
  • Hands-on experience taking LLM systems to production with RAG, retrieval and embedding design, prompt and model versioning, and tool or function calling.
  • Expert-level TypeScript and Python skills, plus experience with REST, OpenAPI, AWS, Docker, CloudFormation or AWS CDK, and CI/CD automation.
  • Experience with AI evaluation, observability, SLI/SLO/SLA practices, distributed tracing, sensitive-data protection, tenant authorization boundaries, and LLM security risks.

Nice to have

  • Experience with advanced LLM orchestration, agentic systems, MCP, multimodal inputs, model optimization, or per-tenant AI cost attribution.
  • Background delivering AI or ML systems in regulated environments such as GxP, 21 CFR Part 11, or SOC 2.
  • Experience in scientific, life sciences, or laboratory data domains.

Culture & Benefits

  • 100% employer-paid benefits for eligible employees and immediate family members.
  • Unlimited paid time off.
  • 401(k) and company-paid life insurance and LTD/STD coverage.
  • Flexible remote working arrangements.
  • Continuous improvement culture with career growth and coaching opportunities.
  • Visa sponsorship is not currently provided.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →