4 часа назад
Principal AI/ML Engineer - AI Safety & Evaluation
225 000 - 245 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal AI/ML Engineer - AI Safety & Evaluation (LLM Safety): Architecting model-level defenses and evaluation infrastructure for foundation models to prevent jailbreaks, prompt injection, toxic outputs, and policy violations with an accent on scalable machine learning systems, adversarial testing, and safety interventions. Focus on designing safety-focused fine-tuning, prompt shielding, post-processing, and human-in-the-loop pipelines while translating emerging threats into measurable safeguards for production.
Location: San Jose, California; hybrid
Targeted compensation: $225,000–$245,000 per year
Company
provides the organizational setting for work focused on making AI systems safe, aligned, and robust against misuse.
What you will do
- Architect and lead model-level defenses against jailbreaks, prompt injection, custom policy violations, and harmful outputs.
- Define adversarial testing, evaluation, and stress-testing strategies to identify safety weaknesses before deployment.
- Set technical direction for safety-focused fine-tuning, prompt shielding, and post-processing mitigation techniques.
- Collaborate with red teamers and researchers to turn emerging threats into measurable evaluations and system safeguards.
- Scale human-in-the-loop pipelines for detecting toxic, biased, and non-compliant outputs.
- Apply current LLM safety research and adversarial trends to production defenses.
Requirements
- 7+ years of experience in applied machine learning, AI infrastructure, or safety-critical systems, including 3+ years in a senior or staff-level technical leadership role.
- Deep understanding of transformer-based architectures and experience building or evaluating LLM safety interventions.
- Expertise in analyzing and addressing adversarial behavior, edge-case failures, and misuse scenarios.
- Ability to guide long-term technical strategy, influence organizational direction, and mentor cross-functional teams.
- Strong written and verbal communication skills with experience influencing technical direction at the organizational or platform level.
- Bachelor’s, Master’s, or PhD in Computer Science, Machine Learning, or a related field.
Nice to have
- Experience with reinforcement learning from human feedback, adversarial training, or safety fine-tuning at scale.
- Hands-on experience designing prompt-level defenses, content filtering systems, or jailbreak prevention mechanisms.
- Contributions to AI safety research, industry standards, or open-source tools for robustness, alignment, or evaluation.
- Familiarity with model governance frameworks, safety policies, model cards, red teaming protocols, or risk classification.
Culture & Benefits
- Work at the intersection of applied research, systems engineering, model behavior, and product safety.
- Build safety systems intended for real-world production deployment.
- is an equal opportunity employer and a VEVRAA federal subcontractor.
Hiring process
- Live interviews are designed to assess the candidate’s own skills and thinking.
- AI or recording tools are not permitted during live interviews unless explicitly invited or approved in advance as a reasonable accommodation.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
ML Research Engineer (AI)
220 000 - 300 000$
1 день назад
Distinguished Principal Engineer - AI Science
225 000 - 250 000$
3 дня назад
AI Applications Engineer (LLM)
100 000 - 175 000$
7 дней назад
Artificial Intelligence and Machine Learning Engineer
99 000 - 225 000$
6 дней назад
Senior Machine Learning Engineer (AI)
185 000 - 230 000$
1 час назад
Principal Engineer, AI Applied Research (AI)
255 850 - 361 200$