5 часов назад
Manager - Data Aggregation
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Manager - Data Aggregation (Python/PySpark): Building secure, compliant pipelines and anonymized datasets from patient-level healthcare data with an accent on tokenization, expert determination, HIPAA compliance, and data quality. Focus on designing end-to-end aggregation solutions, managing parallel client projects, and coordinating delivery with U.S. stakeholders during overlapping working hours.
Location: Pune, India; hybrid workplace. The role requires 3–4 hours of overlap with U.S. working hours to support U.S. clients.
Company
helps life sciences companies improve commercialization through strategy, analytics, artificial intelligence, and technology solutions. The company is headquartered in Evanston, Illinois, and operates 10 global offices.
What you will do
- Collect, clean, standardize, and aggregate patient-level data from EHRs, laboratory systems, and external databases.
- Build and maintain secure data pipelines and repositories for sensitive healthcare information.
- Apply tokenization and support expert determination processes to de-identify data and reduce re-identification risks.
- Generate anonymized datasets for research, analytics, and external data-sharing initiatives while monitoring quality and documenting data provenance.
- Design end-to-end data aggregation solutions using Python and PySpark.
- Lead implementation activities across multiple projects, including client communication, meetings, agendas, recommendations, and delivery management.
Requirements
- Engineering or master’s degree in computer science or a relevant concentration, with strong academic performance.
- 6+ years of relevant consulting-industry experience.
- Strong knowledge of data management, data modeling, data analytics, and U.S. pharmaceutical datasets.
- Experience with tokenization methods, expert determination, HISEC or other high-security frameworks, and sensitive-data handling.
- Strong Python skills and working knowledge of PySpark; familiarity with Snowflake, Redshift, Airflow, and Databricks workflows.
- Ability to support U.S. clients during U.S. working hours with 3–4 hours of overlap, alongside excellent communication, organization, and time-management skills.
Culture & Benefits
- Collaborative, values-driven environment focused on technical excellence and analytical rigor.
- Opportunities to contribute to AI innovation and data-first solutions in life sciences.
- Work in an international matrix environment with cross-functional collaboration.
- Focus on personal growth and client impact.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Senior Cloud Data Engineer (AWS)
3 дня назад
Senior Data Engineer (Snowflake)
155 000 - 165 000$
Snowflake
2 дня назад
Manager/Sr. Manager, Finance Analytics (AI)
200 000 - 262 500$
3 дня назад
Staff Analytics Engineer (Snowflake)
180 000 - 221 000$
14 часов назад
Senior Data Engineer (Azure)
133 000 - 150 000$
21 час назад