Назад
Company hidden
3 дня назад

Data Engineer, AI Enablement

84 500 - 162 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Engineer, AI Enablement (AI/Data Engineering): Building reliable, curated, and AI-ready R&D data products and pipelines for analytics, knowledge graphs, machine learning, and retrieval use cases with an accent on data curation, governance, metadata, lineage, and quality. Focus on preparing data for RAG, embeddings, vector search, and knowledge graph publication while coordinating technical delivery and maintaining production reliability.

Location: North Chicago, Illinois, United States

Salary: $84,500–$162,000 annual base pay

Company

hirify.global develops medicines and healthcare solutions across immunology, oncology, neuroscience, and aesthetics.

What you will do

  • Design, build, and operate scalable data pipelines and curated, reusable data products for R&D.
  • Integrate, normalize, model, document, and publish data from databases, applications, APIs, licensed sources, and other systems.
  • Prepare data and documents for analytics, machine learning, RAG, embeddings, vector search, and ARCH knowledge graph use cases.
  • Apply data quality, metadata, lineage, governance, access control, privacy, licensing, and documentation practices.
  • Collaborate with data scientists, ML engineers, platform teams, architects, data owners, vendors, contractors, and R&D stakeholders.
  • Monitor pipeline performance, data freshness, costs, failures, and delivery issues while improving automation and reusable engineering patterns.

Requirements

  • Bachelor’s degree with 5 years of experience, or master’s degree with 4 years of experience, in information technology, data engineering, data management, analytics, life sciences, or a related field.
  • Hands-on experience with production data pipelines and curated data products using SQL, Python, ETL/ELT patterns, and orchestration tools such as Airflow.
  • Knowledge of modern data platforms, data integration, data warehousing or lakehouse patterns, distributed SQL or big data environments, cloud infrastructure, and analytics enablement.
  • Experience with data cleaning, standardization, enrichment, metadata organization, data modeling, governance, lineage, documentation, and support for embedding or vector-search workflows.
  • Ability to work autonomously, manage concurrent priorities, communicate clearly, resolve issues, and coordinate contributors without direct reporting authority.
  • Experience providing technical guidance and reviewing deliverables from contracted, vendor, or managed-services engineers.

Nice to have

  • Pharmaceutical, healthcare, research, clinical, scientific, or other life sciences data experience.
  • Experience with graph databases, knowledge graphs, ontologies, semantic data, or linked-data concepts.
  • Experience with AWS, Databricks, Spark, Snowflake, Neo4j, lakehouse, or other modern data platforms.
  • Exposure to machine learning, RAG, embeddings, vector databases, AI search, or AI-ready data products.
  • Familiarity with Agile practices and Jira.

Culture & Benefits

  • Work within hirify.global’s R&D data transformation organization.
  • Benefits may include paid vacation, holidays, sick leave, medical, dental, vision insurance, and a 401(k) for eligible employees.
  • Eligibility to participate in short-term incentive programs.
  • Work is governed by GxP, data security, software development lifecycle, data governance, and regulatory standards.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →