2 дня назад
Data Engineer (Databricks)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Data Engineer (Databricks): Building and maintaining batch and streaming data pipelines and governed data products with an accent on Databricks, Python, SQL, Apache Spark, and Delta Lake. Focus on implementing data-quality rules, configuring production workflows, troubleshooting pipeline performance, and supporting secure U.S. public-sector data environments.
Location: Remote within the United States; up to 5% travel. U.S. Citizenship is required, and the role requires the ability to obtain a U.S. Public Trust suitability determination.
Company
is a purpose-driven consulting company delivering mission-focused technology and data solutions for clients.
What you will do
- Develop, test, deploy, and maintain batch and streaming data pipelines using Databricks, Python, SQL, Apache Spark, and Delta Lake.
- Build ingestion, transformation, validation, and publishing processes for governed data products and analytics-ready datasets.
- Implement data models, data-quality rules, metadata, documentation, and source-to-target mappings.
- Configure and maintain Databricks notebooks, workflows, jobs, compute resources, and deployment artifacts.
- Monitor pipeline performance, job execution, data quality, and platform alerts while troubleshooting production incidents.
- Collaborate with architects, analysts, data owners, and engineers on requirements, dependencies, releases, and continuous improvements.
Requirements
- Bachelor’s degree in a technical discipline and at least three years of relevant experience in data engineering, software engineering, analytics engineering, or a related field.
- Proficiency in Python and SQL, including ETL/ELT development, debugging, and maintenance.
- At least one year of hands-on experience with Databricks, Apache Spark, or a comparable cloud data-engineering platform.
- Experience with structured or unstructured data, data validation, production support, and data-processing code.
- Databricks Certified Data Engineer Associate certification is preferred; candidates without it must be willing to obtain it within three months of starting.
- U.S. Citizenship and the ability to obtain a U.S. Public Trust suitability determination are required.
Nice to have
- Experience with Delta Lake, Auto Loader, Databricks SQL, Lakeflow Jobs, Unity Catalog, or streaming data pipelines.
- Familiarity with Git, code reviews, automated testing, CI/CD, and Agile delivery.
- Experience with data governance, metadata, data-quality checks, lineage, and access controls.
- Experience with AWS, Azure, Google Cloud, regulated data, public-sector environments, or security-sensitive systems.
- Additional Databricks certifications, including Machine Learning Engineer or Generative AI Engineer certifications.
Culture & Benefits
- Inclusive culture centered on passion, solidarity, generosity, curiosity, boldness, and authenticity.
- Generous paid time off and federal holidays.
- Parental leave and comprehensive medical, dental, vision, life, and disability coverage.
- 401(k) with company match and FSA/HSA options.
- Commuter benefits and a remote work arrangement with limited travel.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →