Назад
Company hidden
2 дня назад

Data Engineer (Databricks)

Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Engineer (Databricks): Building and maintaining batch and streaming data pipelines and governed data products with an accent on Databricks, Python, SQL, Apache Spark, and Delta Lake. Focus on implementing data-quality rules, configuring production workflows, troubleshooting pipeline performance, and supporting secure U.S. public-sector data environments.

Location: Remote within the United States; up to 5% travel. U.S. Citizenship is required, and the role requires the ability to obtain a U.S. Public Trust suitability determination.

Company

hirify.global is a purpose-driven consulting company delivering mission-focused technology and data solutions for clients.

What you will do

  • Develop, test, deploy, and maintain batch and streaming data pipelines using Databricks, Python, SQL, Apache Spark, and Delta Lake.
  • Build ingestion, transformation, validation, and publishing processes for governed data products and analytics-ready datasets.
  • Implement data models, data-quality rules, metadata, documentation, and source-to-target mappings.
  • Configure and maintain Databricks notebooks, workflows, jobs, compute resources, and deployment artifacts.
  • Monitor pipeline performance, job execution, data quality, and platform alerts while troubleshooting production incidents.
  • Collaborate with architects, analysts, data owners, and engineers on requirements, dependencies, releases, and continuous improvements.

Requirements

  • Bachelor’s degree in a technical discipline and at least three years of relevant experience in data engineering, software engineering, analytics engineering, or a related field.
  • Proficiency in Python and SQL, including ETL/ELT development, debugging, and maintenance.
  • At least one year of hands-on experience with Databricks, Apache Spark, or a comparable cloud data-engineering platform.
  • Experience with structured or unstructured data, data validation, production support, and data-processing code.
  • Databricks Certified Data Engineer Associate certification is preferred; candidates without it must be willing to obtain it within three months of starting.
  • U.S. Citizenship and the ability to obtain a U.S. Public Trust suitability determination are required.

Nice to have

  • Experience with Delta Lake, Auto Loader, Databricks SQL, Lakeflow Jobs, Unity Catalog, or streaming data pipelines.
  • Familiarity with Git, code reviews, automated testing, CI/CD, and Agile delivery.
  • Experience with data governance, metadata, data-quality checks, lineage, and access controls.
  • Experience with AWS, Azure, Google Cloud, regulated data, public-sector environments, or security-sensitive systems.
  • Additional Databricks certifications, including Machine Learning Engineer or Generative AI Engineer certifications.

Culture & Benefits

  • Inclusive culture centered on passion, solidarity, generosity, curiosity, boldness, and authenticity.
  • Generous paid time off and federal holidays.
  • Parental leave and comprehensive medical, dental, vision, life, and disability coverage.
  • 401(k) with company match and FSA/HSA options.
  • Commuter benefits and a remote work arrangement with limited travel.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →