Назад
5 дней назад

Safeguards Enforcement Analyst (AI)

245 000 - 330 000$
Формат работы
hybrid
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Safeguards Enforcement Analyst (AI): Designing and operating automated enforcement systems, review workflows, and evaluations to detect and mitigate misuse of AI systems involving conventional weapons and dangerous technology with an accent on policy enforcement, threat intelligence, and model safety. Focus on analyzing flagged content, identifying emerging misuse patterns and attack vectors, and translating technical evidence into accurate enforcement decisions.

Location: New York City, NY; remote-friendly with travel required; San Francisco, CA; or Washington, DC. Staff are expected to work from one of the offices at least 25% of the time.

Annual salary: $245,000–$330,000 USD.

Company

Anthropic develops reliable, interpretable, and steerable AI systems intended to be safe and beneficial for users and society.

What you will do

  • Design and scale automated enforcement systems and content-review workflows while maintaining high accuracy.
  • Develop and maintain evaluations that measure model performance, identify regressions, and inform policy and model improvements.
  • Partner with Engineering and Data Science to improve detection and automated enforcement systems.
  • Review flagged content, make enforcement decisions, and identify policy gaps and sophisticated misuse attempts.
  • Develop enforcement guidelines and reviewer documentation for consistent decisions across diverse content.
  • Monitor emerging weapons trends, regulatory changes, misuse patterns, attack vectors, and coordinated violent activity.

Requirements

  • Deep applied expertise in weapons systems and the ability to translate complex technical evidence into enforcement decisions.
  • Experience in policy enforcement, threat intelligence, counterterrorism, government, or a related field involving harmful content, dangerous technology, or physical-harm facilitation.
  • Experience creating and scaling policy-enforcement or content-review workflows.
  • Proficiency in SQL or other data-analysis tools for analyzing large datasets and monitoring workflow health.
  • Experience identifying emerging risks and threat actors and communicating findings to Product, Policy, Engineering, and Legal stakeholders.
  • Experience with generative AI products, effective prompting for content review, and implementing product policies at scale.

Nice to have

  • Expertise in conventional weapons, dangerous technology, autonomous systems, or critical infrastructure protection.
  • Experience developing evaluations or red-teaming AI systems, including harmful-content and policy-enforcement use cases.
  • Experience with threat-actor profiling, MITRE ATT&CK, OSINT, and cross-platform threat analysis.
  • Proficiency in Python for data analysis and workflow automation.
  • Background in law enforcement, national security, defense, counterterrorism, or relevant regulation.

Culture & Benefits

  • Collaborative work across research, engineering, policy, and business disciplines.
  • Competitive compensation with optional equity donation matching.
  • Generous vacation and parental leave.
  • Flexible working hours and office collaboration.
  • Visa sponsorship is available, subject to role and candidate eligibility.

Hiring process

  • A bachelor’s degree or equivalent combination of education, training, and experience is required.
  • The required field of study must be relevant to the role through coursework, training, or professional experience.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →