4 месяца назад
Data Engineer (Python/BigQuery)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Data Engineer (Python/BigQuery): Building and operating large-scale batch and streaming data infrastructure for analytics, machine learning, and CDC delivery with an accent on reliable ingestion, cloud data platforms, and governance. Focus on designing resilient pipelines, implementing CI/CD and data-quality controls, and developing internal tools with LLM-based agents.
Location: Hybrid in Almaty, Kazakhstan
Company
is a global technology company providing ride-hailing, freight, delivery, intercity travel, and financial services across 47 countries.
What you will do
- Build and operate batch and streaming ingestion into layered BigQuery data warehouses using Airflow, Debezium, Kafka, Pub/Sub, and Dataflow.
- Integrate marketing platforms, payment providers, S3 buckets, and third-party APIs, including schema contracts, backfills, and reconciliation.
- Develop Python-based platform tools, Airflow operators, connectors, Kafka Connect services, Cloud Functions, and API integrations.
- Build BigQuery CI/CD and change-management tooling with testing, migrations, sandbox validation, backup, and rollback workflows.
- Improve pipeline reliability through idempotency, deduplication, late-data handling, replay, monitoring, alerting, and automated testing.
- Develop governance and internal data tools, support analysts and business teams, and contribute to system architecture.
Requirements
- Strong practical Python experience with clean, structured, and tested code for services, tooling, and data pipelines.
- Solid software design skills, including OOP, modularity, and design patterns.
- Experience operating cloud services with CI/CD, containerization, monitoring, and alerting in GCP, AWS, or a similar environment.
- Familiarity with Kubernetes and Terraform, plus hands-on experience with BigQuery or another cloud data warehouse and confident SQL skills.
- Clear communication with non-engineering stakeholders and the ability to take ownership of technologies and services.
- Availability for hybrid work in Almaty, Kazakhstan.
Nice to have
- Advanced SQL, analytical data modeling, and experience optimizing warehouse performance and costs.
- Experience with CDC pipelines, Debezium, schema evolution, backfills, reconciliation, Flink, Apache Beam, or Dataflow.
- Knowledge of ITGC/SOX compliance, Databricks Unity Catalog, OpenMetadata, or Dataplex.
- Interest in building LLM-based agents and AI tooling for data.
Culture & Benefits
- Mentoring, career consulting, and professional learning programs.
- International collaboration and participation in the Global Talent Exchange Program.
- Company-wide challenges, awards, sports activities, volunteering, and social-impact projects.
- Language courses and internal speaking clubs.
- Benefits may vary depending on location.
Hiring process
- Application review.
- Offer.
- Hiring.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
COMTEK
12 дней назад
Data Analytics Engineer (AI)
4 000 - 5 500$
Страховая компания Freedom Insurance
11 дней назад
Data Engineer
11 дней назад
Senior Snowflake Developer with AWS
Республиканский центр электронного здравоохранения
13 дней назад
Инженер по данным (Python)
600 000KZT
Solva
5 дней назад
Chief Data Officer (Fintech)
9 дней назад