обновлено 2 дня назад
Hadoop Solutions Developer (Big Data)
100 000 - 150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Hadoop Solutions Developer (Big Data): Building and operating large-scale Hadoop data processing pipelines and analytics platforms with an accent on Spark-based ETL/ELT, streaming ingestion, and distributed data storage. Focus on tuning Spark and MapReduce workloads, designing reliable lakehouse layouts, automating orchestration, and implementing monitoring, governance, and data quality controls.
Location: 100% remote within the United States
Salary: $100,000–$150,000 annually
Company
is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
What you will do
- Design, develop, and operate end-to-end Hadoop pipelines using relational, file-based, streaming, and API-driven data sources.
- Build ETL/ELT workflows with Spark, Hive, Pig, and Sqoop, ensuring data quality, recoverability, and reliable error handling.
- Develop high-throughput streaming pipelines with Kafka, Spark Streaming, or Flink.
- Optimize Spark and MapReduce jobs, data models, storage layouts, and lakehouse formats for performance and cost efficiency.
- Implement governance, lineage, quality controls, monitoring, alerting, logging, and operational runbooks.
- Partner with data scientists and analysts, review existing architectures, and mentor junior engineers.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
- At least 5 years of professional experience designing and operating Hadoop big-data pipelines.
- Production expertise with Apache Spark using Scala, Python, or Java, plus strong SQL skills.
- Experience with Hive, HDFS, Sqoop, HBase, Kafka, Spark Streaming, or Flink.
- Experience with workflow orchestration tools such as Airflow or Oozie and relational and NoSQL data stores.
- Understanding of distributed systems, strong Python or Shell scripting, troubleshooting, debugging, and documentation skills.
Nice to have
- Experience with AWS EMR, Azure HDInsight, or Databricks.
- Familiarity with Delta, Iceberg, or Hudi lakehouse formats.
- Exposure to Apache Atlas or Collibra, Kubernetes-based data platforms, CI/CD, or infrastructure as code.
Culture & Benefits
- Full-time direct W2 employment.
- Career growth opportunities within an established organization.
- Opportunity to work on cloud, AI, data, and enterprise solutions.
Hiring process
- Submit a resume for consideration.
- U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply.
- New H-1B visa petitions cannot be sponsored for this position.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Plaid
3 дня назад
Senior Data Engineer
191 000 - 239 000$
4 дня назад
Senior Data Engineer (Databricks)
115 000 - 165 000$
4 дня назад
Data Engineer
95 553 - 165 000$
6 дней назад
Senior People Data and AI Engineer
142 600 - 257 600$
6 дней назад
Data Engineer / Technical Lead (Apache)
250 000 - 270 000₽
3 дня назад
Senior Field Specialist - Data Engineering & Streaming (Pre-Sales)
166 000 - 230 000$