Назад
8 дней назад

Safeguards Enforcement Lead, User Well-Being

285 000 - 330 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
lead
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Safeguards Enforcement Lead, User Well-Being (AI Safety): Managing child safety, mental health, abuse and exploitation, and age assurance enforcement workflows with an accent on content review operations, policy enforcement, and automated detection systems. Focus on scaling accurate review processes, optimizing AI-driven enforcement, identifying emerging misuse patterns, and coordinating legally required reporting.

Location: New York City, San Francisco, or Washington, DC; hybrid attendance required at least 25% of the time

Annual salary: $285,000–$330,000 USD

Company

Anthropic builds reliable, interpretable, and steerable AI systems designed to be safe and beneficial for users and society.

What you will do

  • Manage individual contributors across child safety, mental health, abuse and exploitation, and age assurance policy areas.
  • Coordinate content review partners, including onboarding, training, quality assurance, and relationship management.
  • Design and scale enforcement workflows, documentation, decision trees, and review guidelines while maintaining accuracy and consistency.
  • Partner with Engineering and Data Science to improve detection models and automated enforcement systems.
  • Identify misuse trends and communicate emerging risks to Product, Policy, Engineering, Legal, and Trust & Safety stakeholders.
  • Coordinate reporting obligations to external bodies such as NCMEC in accordance with applicable law and policy.

Requirements

  • Experience managing teams in user well-being, trust and safety, content moderation operations, or policy enforcement.
  • Direct experience with child safety, mental health, abuse and exploitation, and age assurance harms.
  • Experience managing or coordinating content review operations, quality assurance, and workflow scaling.
  • Proficiency in SQL and/or other data analysis tools for monitoring review queues, workflow health, and enforcement trends.
  • Experience identifying emerging risks and communicating findings cross-functionally.
  • Bachelor’s degree or equivalent education, training, or professional experience.

Nice to have

  • Expertise in child sexual exploitation and abuse, online child protection, mental wellness, or age assurance.
  • Experience with NCMEC, IWF, or equivalent reporting bodies and relevant regulatory frameworks such as KOSA or COPPA.
  • Experience with generative AI products, trauma-informed reviewer support, Python, hash-matching technologies, or age assurance systems.

Culture & Benefits

  • Collaborative environment focused on large-scale AI safety research and trustworthy AI.
  • Competitive compensation with optional equity donation matching.
  • Generous vacation and parental leave.
  • Flexible working hours and access to wellness resources for exposure to psychologically disturbing content.
  • Visa sponsorship is available, subject to role and candidate eligibility.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →