Назад
Company hidden
3 часа назад

Senior Data Scientist (Biologics Discovery)

55 400 - 87 860
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US/Spain
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Data Scientist (Biologics Discovery) (Python/PyTorch/SQL): Building model-ready datasets, applied ML models, and evaluation frameworks from biologics assay, biophysical, antibody, and protein sequence data with an accent on robust featurization, reproducibility, and trustworthy model evaluation. Focus on detecting leakage and distribution shift, enabling molecular property modeling, and translating experimental data into actionable insights for biologics discovery.

Location: Office-based in Spring House, PA; Titusville, NJ; Raritan, NJ; or Madrid, Spain. No remote option.

Base pay: €55,400–€87,860 annually, depending on pay grade and location.

Company

Johnson & Johnson develops healthcare, Innovative Medicine, and MedTech solutions focused on preventing and treating complex diseases.

What you will do

  • Develop featurization approaches and model-ready datasets from antibody and protein sequences, constructs, assays, and biophysical data.
  • Define features, labels, and aggregation levels with data engineers while preserving important raw representations.
  • Curate, document, and version datasets to support reproducible and traceable modeling.
  • Build and evaluate applied ML models and evaluation frameworks for biologics discovery data.
  • Collaborate with In Silico Discovery, experimental scientists, ontology specialists, and MLOps colleagues across the design-make-test-learn cycle.
  • Support responsible AI, consistent data semantics, and reliable movement of models from development into use.

Requirements

  • Master’s or Ph.D. in Computer Science, Machine Learning, Computational Biology, Bioinformatics, Statistics, or a related field.
  • At least 2 years of applied ML experience, including model development, evaluation, and dataset curation involving complex scientific or biomedical data.
  • Strong proficiency in Python, PyTorch or scikit-learn, and SQL.
  • Experience transforming heterogeneous experimental data into robust features and training sets, with exposure to cloud training and data infrastructure.
  • Understanding of evaluation, validation, data leakage, and distribution shift risks.
  • Ability to collaborate with experimental scientists and modeling partners in a matrixed R&D environment.

Nice to have

  • Experience with biologics, antibody or protein sequence models, or protein language models.
  • Experience with active learning, Bayesian optimization, or sequence-based generative models for molecular design.
  • Familiarity with biophysical data, assay data, and developability endpoints.
  • Experience with MLOps, experiment tracking, and model monitoring.
  • Knowledge of ontologies or knowledge graphs for data reuse and AI-ready datasets.

Culture & Benefits

  • Inclusive work environment based on individual consideration, diversity, and dignity.
  • Annual bonus with a target determined by pay grade and location.
  • Vacation, parental leave of at least 12 weeks, bereavement, caregiver, and volunteer leave.
  • Well-being reimbursement and financial, physical, and mental health programs.
  • Recognition awards and access to insurance plans subject to applicable plan terms and location eligibility.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →