обновлено 24 дня назад
Senior Data Engineer (GCP)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Data Engineer (GCP): Building secure, automated, and scalable data-processing pipelines for high-volume advertising and analytics systems with an accent on Python, GCP, data governance, and production reliability. Focus on architecting Apache Druid and Imply platforms, optimizing distributed processing with Spark, and designing event-driven pipelines that handle petabyte-scale data and high request volumes.
Location: Mumbai, India; hybrid with 3 days onsite. Shift: 12 pm - 8 pm IST.
Company
is a location-powered advertising and analytics platform that helps brands and media agencies understand audiences and deliver targeted campaigns without relying on personal data.
What you will do
- Design, build, monitor, and support large-scale data-processing pipelines.
- Develop secure, automated, scalable production pipeline code primarily in Python.
- Maximize cloud-compute efficiency and ensure data quality, governance, recovery, and operational readiness.
- Explore new data streams to support commercial and technical growth.
- Support, mentor, and pair with other engineers to improve team capabilities.
- Collaborate closely with Product to deliver fast, robust features for customers.
Requirements
- 8+ years of experience delivering robust, performant data pipelines under direct SLAs and commercial constraints.
- Experience architecting, developing, and maintaining Apache Druid and Imply platforms, including DevOps practices and large-scale re-architecture.
- Advanced experience building GCP pipelines with native services and Apache Airflow.
- Strong Python skills for data and computational tasks, including cleansing, validation, and composition.
- Hands-on experience with streaming data, relational and non-relational databases, and distributed processing technologies such as Spark.
- Strong knowledge of data-science libraries, cloud services, server-side Linux, documentation, testing, and solution assurance.
Nice to have
- Experience optimizing code and configuration in Spark, Hive, or similar tools.
- Experience with relational database partitioning and indexing.
- Knowledge of AWS Athena or Google BigQuery for data-centric problems.
- Experience applying statistical techniques and complex algorithms to large data structures.
- Experience with Python notebooks such as Jupyter, Zeppelin, or Google Datalab.
Culture & Benefits
- Work in a growing engineering team focused on ambitious scale and intelligence goals.
- Lean development environment with significant autonomy and emphasis on critical thinking.
- Inclusive workplace committed to diversity, belonging, equal opportunity, and employee wellbeing.
- Company values include innovation, ownership, client focus, teamwork, and action-oriented delivery.
- operates across more than 40 offices on five continents and works with global brands and media agencies.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →