Назад
Company hidden
2 часа назад

Senior Data Scientist (AI)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Data Scientist (AI) (Python/SQL, Data Engineering): Developing and validating interpretable analytical signals and data pipelines for a secure Federal analytics and AI modernization platform with an accent on structured and unstructured data, machine learning, NLP, and traceable AI-assisted analytics. Focus on building reproducible ETL/ELT workflows, preparing embeddings and retrieval datasets, implementing data quality and provenance controls, and supporting secure on-premises and cloud environments.

Location: Bethesda, MD, United States. The role supports a Federal initiative and requires U.S. citizenship and the ability to obtain and maintain Top Secret eligibility.

Company

hirify.global is a consulting and information technology solutions provider delivering software, systems engineering, information management, analytics, visualization, and management consulting services to public sector organizations.

What you will do

  • Design, develop, and maintain data pipelines and ETL/ELT workflows for machine learning, NLP, information retrieval, and advanced analytics.
  • Prepare, transform, curate, and validate structured and unstructured grants and financial data for AI/ML applications.
  • Build features, process documents, generate embeddings, and prepare datasets for search, RAG, visualization, and AI-assisted decision support.
  • Implement data quality, validation, provenance, lineage, versioning, metadata, and source-traceability controls.
  • Troubleshoot and optimize data integrations and processing workflows while collaborating with data engineering, data science, software engineering, and AI/ML teams.
  • Document data architectures, schemas, transformations, dependencies, embeddings, and operational procedures for secure Federal environments.

Requirements

  • Experience building, integrating, and operating data pipelines for machine learning, NLP, information retrieval, or advanced analytics.
  • Hands-on experience with feature engineering, document processing, embeddings, curated datasets, reproducible workflows, and structured and unstructured data.
  • Strong proficiency in Python and SQL.
  • Experience implementing data quality controls, validation, provenance, metadata, versioning, lineage, and source traceability.
  • Strong technical communication, problem-solving, documentation, and cross-functional collaboration skills.
  • U.S. citizenship and ability to obtain and maintain Top Secret eligibility required; active Top Secret clearance highly preferred.

Nice to have

  • Experience with Federal financial, grants, budget, payment, award oversight, or financial reporting data.
  • Experience with secure Federal, on-premises, cloud, or hybrid environments.
  • Experience with Linux, PostgreSQL, Microsoft SQL Server, Java, .NET, Jenkins, CI/CD, Azure DevOps, Kubernetes, or Rancher.
  • Familiarity with open-source AI/ML models, RAG, vector databases, semantic search, embeddings pipelines, and Federal data governance requirements.

Culture & Benefits

  • Work with multidisciplinary professionals supporting health, civilian, and national security missions.
  • Contribute to a secure, scalable platform that supports examiner search, review, monitoring, and decision support.
  • Employer-paid health care.
  • Training and development funds.
  • 401(k) match and opportunities for bonuses.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →