Назад
Company hidden
5 дней назад

Data Engineer – Legacy Systems & AI Workflows

89 886 - 175 444$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Engineer – Legacy Systems & AI Workflows (ETL/ELT, SOAP, AI): Building secure batch and streaming data pipelines that integrate legacy and modern sources into cloud data platforms with an accent on data quality, governance, provenance, and access controls. Focus on onboarding SOAP-based legacy systems, troubleshooting distributed pipelines, and developing preprocessing and metadata-enrichment workflows for AI/ML capabilities.

Location: McLean, Virginia, United States. The role requires U.S. citizenship and the ability to obtain and maintain a Public Trust Clearance. Preference is given to candidates local to the Washington, DC metro area.

Salary: $89,886–$175,444 per year.

Company

hirify.global is a technology integrator and solution provider delivering secure technology solutions for federal agencies.

What you will do

  • Design, develop, and operate batch and streaming ETL/ELT pipelines for structured and semi-structured data.
  • Onboard legacy data sources by reviewing schemas, mappings, interfaces, validation rules, ownership, and operational procedures.
  • Build automated data-quality checks, monitoring, alerting, and validation processes.
  • Automate deployments and pipeline operations using SOAP APIs, workflow orchestration, and environment-specific configurations.
  • Troubleshoot distributed data systems using logs, metrics, lineage, and operational signals.
  • Collaborate with software, cloud, data science, security, architecture, and customer teams to deliver measurable data products.

Requirements

  • Bachelor’s degree in computer science, computer engineering, mathematics, statistics, or a related technical field.
  • At least 5 years of professional data engineering experience, including production pipeline development and operations.
  • Programming experience in Python, Java, R, or another data engineering language.
  • Experience with legacy systems using SOAP APIs, REST APIs, and structured formats including JSON, XML, CSV, and Parquet.
  • Experience with data-quality frameworks, monitoring, alerting, security, access controls, encryption, auditability, and governance in regulated or restricted environments.
  • U.S. citizenship and eligibility to obtain and maintain a Public Trust Clearance are required.

Nice to have

  • Experience building preprocessing, chunking, filtering, metadata-enrichment, or evaluation pipelines for LLM and AI/ML workflows.
  • Experience with segmented networks, limited connectivity, strict change control, or formal authorization requirements.
  • Strong SQL skills, including query optimization, data modeling, joins, window functions, and large-dataset analysis.
  • Experience with Jupyter Notebook, NumPy, pandas, AWS, C3.ai, or relevant cloud and data engineering certifications.
  • Active Public Trust Clearance and proximity to the Washington, DC metro area.

Culture & Benefits

  • Medical, dental, and vision plans.
  • Flexible Spending Account and a 4% 401(k) match with immediate vesting.
  • Paid time off, tuition reimbursement, certification programs, and professional development.
  • Flexible work schedule.
  • On-site gym and childcare option.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →