Назад
Company hidden
4 часа назад

Staff Data Engineer (AI)

Формат работы
remote (только Canada)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff Data Engineer (AI): Designing and scaling data infrastructure, ETL/ELT pipelines, and lakehouse architecture for analytics, machine learning, and business intelligence with an accent on Databricks, Spark optimization, and reliable data platforms. Focus on tuning distributed data workloads, building observability and data quality systems, and driving architectural standards across high-scale data engineering.

Location: Canada — Remote

Company

hirify.global provides software supply chain security, SBOM management, open source dependency management, and repository solutions for enterprises and software developers.

What you will do

  • Design, build, and maintain scalable data pipelines and ETL/ELT processes.
  • Architect and optimize data models, storage solutions, and next-generation data lakehouse architecture.
  • Own and evolve data platform components using Databricks and Spark.
  • Deliver trusted datasets in collaboration with data scientists, analysts, engineers, and business stakeholders.
  • Implement observability, alerting, data quality monitoring, documentation, testing, and CI/CD practices.
  • Drive long-term architectural vision and mentor engineers on data engineering best practices.

Requirements

  • 8+ years of experience as a Data Engineer or in a similar backend engineering role.
  • Bachelor’s degree in Computer Science, Engineering, or a related technical field.
  • Strong programming skills in Python, Scala, or Java.
  • Hands-on experience with distributed data systems such as Spark or Kafka, and production ETL/ELT pipelines.
  • Experience tuning Spark jobs, optimizing joins, and managing Delta Lake architecture for batch and streaming data.
  • Advanced SQL and NoSQL querying skills, including query performance optimization and data modeling.

Nice to have

  • Experience with software supply chain, cybersecurity, or large-scale software ecosystem data.
  • Familiarity with Airflow, Dagster, or similar workflow orchestration tools.
  • Experience with AWS and modern table formats such as Delta Lake, Apache Iceberg, or Apache Hudi.
  • Experience with data observability, lineage, governance, automated data quality, and real-time data lake architectures.

Culture & Benefits

  • Collaborative culture focused on learning, autonomy, and impact.
  • Access to open-source and cloud-native technologies.
  • Diversity and inclusion working groups.
  • Parental leave policy.
  • Paid volunteer time off.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →