Назад
7 дней назад

Staff+ Software Engineer, Account Creation (AI Safety)

320 000 - 485 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff+ Software Engineer, Account Creation (AI Safety): Building monitoring systems, abuse detection infrastructure, and multi-layered defenses for AI systems with an accent on trust and safety, real-time enforcement, and scalable oversight. Focus on detecting unwanted model behaviors, surfacing abuse patterns for model hardening, and designing internal tooling for analysts and research teams.

Location: San Francisco, CA; New York City, NY; or Seattle, WA. Hybrid work is required, with staff expected to be in one of the offices at least 25% of the time.

Annual salary: $320,000–$485,000 USD

Company

Anthropic builds reliable, interpretable, and steerable AI systems intended to be safe and beneficial for users and society.

What you will do

  • Develop monitoring systems to detect unwanted behavior from API partners and support automated enforcement actions.
  • Build dashboards that surface potential abuse for analyst review.
  • Design abuse detection mechanisms and supporting infrastructure.
  • Identify abuse patterns and share them with research teams to improve model training.
  • Build reliable, multi-layered safety defenses that improve in real time and operate at scale.

Requirements

  • Bachelor’s degree in Computer Science, Software Engineering, or equivalent education and experience.
  • Proficiency in Python and TypeScript.
  • Ability to work across the stack.
  • Strong communication skills and ability to explain complex technical concepts to non-technical stakeholders.
  • Ability to work from and regularly attend an Anthropic office in San Francisco, New York City, or Seattle.

Nice to have

  • 8+ years of software engineering experience.
  • Experience with integrity, spam, fraud, or abuse detection and mitigation.
  • Experience building trust and safety detection mechanisms for AI/ML systems.
  • Experience with prompt engineering, jailbreak attacks, and adversarial inputs.
  • Experience partnering with operational teams to build custom internal tooling.

Culture & Benefits

  • Collaborative work across research, engineering, policy, and business functions.
  • Flexible working hours and office-based collaboration.
  • Generous vacation and parental leave.
  • Competitive compensation, benefits, and optional equity donation matching.
  • Visa sponsorship is available, with immigration counsel provided for supported cases.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →