3 дня назад
Hadoop Solutions Developer (Hadoop)
100 000 - 150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Hadoop Solutions Developer (Hadoop): Building and operating large-scale data processing pipelines and analytics platforms for structured and unstructured data with an accent on Spark, streaming, Hadoop ecosystem technologies, and production reliability. Focus on optimizing distributed workloads, designing lakehouse storage layouts, implementing governance and monitoring, and delivering resilient data platforms for analytics and machine learning.
Location: 100% remote within the continental United States
Salary: $100K–$150K annually
Company
Software development company building scalable, secure, and user-friendly solutions that automate and optimize business operations.
What you will do
- Design, develop, and operate end-to-end big-data pipelines on Hadoop using relational, file-based, streaming, and API-driven data sources.
- Build ETL/ELT workflows with Spark, Hive, Pig, and Sqoop, emphasizing data quality, recoverability, error handling, and idempotency.
- Develop high-throughput streaming pipelines with Kafka, Spark Streaming, or Flink and integrate them with analytical and operational systems.
- Optimize Spark and MapReduce workloads, data models, storage layouts, and lakehouse formats to meet performance and cost requirements.
- Implement monitoring, alerting, logging, governance, lineage, quality controls, and operational documentation for production pipelines.
- Collaborate with data scientists and analysts, review architectures, mentor junior engineers, and contribute to engineering standards.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
- Five or more years of professional experience designing and operating Hadoop-based big-data pipelines.
- Production expertise with Apache Spark and strong experience with Hive, HDFS, Sqoop, HBase, and the wider Hadoop ecosystem.
- Hands-on experience with Kafka, Spark Streaming, or Flink; strong SQL skills; and experience with relational and NoSQL data stores.
- Experience with Airflow or Oozie, distributed systems concepts, Python or Shell scripting, troubleshooting, debugging, and technical documentation.
- Must work as a direct full-time W2 employee in the continental United States; no C2C, 1099, or third-party arrangements.
Nice to have
- Experience with AWS EMR, Azure HDInsight, or Databricks.
- Familiarity with Delta, Iceberg, Hudi, Apache Atlas, or Collibra.
- Experience with Kubernetes-based data platforms, CI/CD, and infrastructure as code.
Culture & Benefits
- Direct employment through an in-house SOW engagement.
- Full-time W2 employment with a remote work arrangement within the continental United States.
- No new H1B sponsorship; qualified candidates with a valid H1B may receive transfer support.
- Technical coding assessment is mandatory.
- Equal opportunity workplace with a focus on diversity, inclusion, and reasonable accommodations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Senior Staff Data Engineer (Fintech)
166 400 - 244 800$
Top Selection
7 дней назад
Data Engineer (MLOps)
280 000 - 300 000₽
7 дней назад
Senior Software Engineer-Bigdata & Hadoop Engineer with Development Experience
110 700 - 171 800$
4 дня назад
Software Engineer - Data (Python)
5 дней назад
Senior Software Engineer (Big Data)
135 000 - 150 000$
Nitka Technologies, Inc.
4 дня назад