1 день назад
Lead Data Engineer (Databricks)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Lead Data Engineer (Databricks) (Databricks, Python, PySpark, SQL): Building scalable ETL/ELT pipelines and production-ready data solutions on the Databricks Lakehouse Platform with an accent on data ingestion, transformation, modeling, and quality. Focus on optimizing Spark workloads, managing Delta Lake tables with Medallion Architecture, and designing reliable workflows for large-scale data processing.
Location: Hybrid in Coimbatore, India
Company
is a growing data product company founded in 2020 that delivers digital solutions primarily for Fortune 500 companies.
What you will do
- Design, develop, maintain, and optimize scalable ETL/ELT pipelines using Databricks, PySpark, and SQL.
- Integrate data from databases, Amazon S3, files, and REST APIs for ingestion and export.
- Build pipelines with Unity Catalog and manage Delta Lake tables using Bronze, Silver, and Gold layers.
- Implement business logic, data transformations, dimensional models, and data warehousing solutions.
- Create, schedule, monitor, and optimize Databricks Jobs and Workflows.
- Ensure data quality and production reliability through validation, error handling, logging, monitoring, and Spark performance tuning.
Requirements
- Strong expertise in Python, PySpark, and advanced SQL.
- Hands-on experience with the Databricks Lakehouse Platform, including Unity Catalog, Delta Lake, Workflows, Jobs, Clusters, Notebooks, and Repos.
- Experience with ETL/ELT development, batch processing, incremental loading, data transformation, and REST API integration.
- Knowledge of dimensional modeling, Star and Snowflake Schemas, Fact and Dimension tables, SCD concepts, and data warehousing.
- Experience with structured and semi-structured data formats including CSV, JSON, Parquet, and Delta.
- Knowledge of partitioning, file optimization, Spark and query performance tuning, Git, and CI/CD best practices.
Nice to have
- Experience with Auto Loader, Spark Declarative Pipelines, Kafka, Airflow, or dbt.
- Databricks certification.
Culture & Benefits
- Competitive salary.
- Strong insurance package.
- Learning and development resources supporting employee growth.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
Senior Snowflake Data Engineer (Snowflake)
3 дня назад
Databricks Genie Consultant / Data Engineer (Databricks)
4 дня назад
Data Engineer-Hyd
6 дней назад
Portfolio Analytics Architect
103 000 - 129 000$
4 дня назад
Portfolio Analytics Architect (Investment Management)
103 000 - 129 000$
5 дней назад