Назад
Company hidden
обновлено 1 день назад

Senior Data Engineer

165 000 - 220 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Релокация
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Data Engineer (PySpark/AWS): Building and operating clinical data services, models, and pipelines that power healthcare analytics and research with an accent on data quality, reliability, and production deployment. Focus on tuning Spark workloads, managing cloud data infrastructure, adapting to upstream changes, and supporting trustworthy downstream datasets.

Location: Hybrid in the New York City, Los Angeles, or San Francisco metro areas; in-office Tuesdays and Thursdays, with additional office presence during onboarding and team onsite weeks. Relocation assistance is available for candidates who do not already reside in the NYC metro area. Hybrid employees may work outside their home office for up to 6 weeks per year.

Salary: $165,000–$220,000 per year plus equity.

Company

hirify.global develops healthcare technology that analyzes medical records to recommend diagnoses and identify patients who may qualify for lifesaving treatment.

What you will do

  • Design, build, and operate data services, models, and pipelines for clinical data.
  • Collect, standardize, and consolidate data for analytics and research initiatives.
  • Deliver reliable datasets to downstream product and analytics workflows.
  • Own data quality and production readiness in collaboration with engineering teams.
  • Monitor, alert, and manage the day-to-day reliability of the data platform and its pipelines.
  • Partner with product stakeholders to deliver analytics and actionable insights.

Requirements

  • Bachelor’s degree in Computer Science, Mathematics, Statistics, or a related field, or equivalent practical experience.
  • 5+ years of experience in data engineering and 3+ years building pipelines with PySpark.
  • 3+ years of experience with public cloud technologies, including AWS tools such as S3, EMR, or Athena.
  • Strong proficiency in Python and SQL, with hands-on experience in data modeling and pipeline design.
  • Practical experience with LLM-assisted development and its capabilities and limitations.
  • US work authorization without visa sponsorship and residence within the New York City, Los Angeles, or San Francisco metro areas are required. Participation in on-call operational support is also required.

Nice to have

  • Experience with Apache Iceberg, Dagster, ClickHouse, PostgreSQL, FastAPI, or Metabase.
  • Experience with healthcare data, HIPAA compliance, data de-identification, or OMOP CDM.
  • Experience supporting production data pipelines, including validation, deployment, and performance evaluation.

Culture & Benefits

  • Equity eligibility.
  • 99% employer-paid medical, dental, and vision benefits plus a One Medical subscription.
  • 18 PTO days per year and a one-week holiday break.
  • Monthly health and wellness budget, company-sponsored retreats, and social events.
  • Sabbatical program and catered lunches at the LA and Manhattan offices.

Hiring process

  • Successful completion of a background check is required.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →