Назад
Company hidden
2 дня назад

Data Engineer, Forward Deployed Engineer (AI)

143 640 - 273 000$
Формат работы
hybrid
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Engineer, Forward Deployed Engineer (AI): Designing and deploying scalable data infrastructure, ETL/ELT pipelines, vector databases, and context engines for AI, machine learning, and autonomous workflows with an accent on production-grade performance, real-time streaming, and customer deployments. Focus on optimizing distributed data systems, building RAG and agent-memory capabilities, implementing observability and data quality controls, and translating field insights into stable platform improvements.

Location: Hybrid role requiring at least 3 days per week onsite in Frisco, Texas; candidates must live within commuting distance. Relocation assistance is not provided.

Salary: $143,640–$273,000 annually in the United States, with location-specific ranges for certain states and cities. Annual bonus eligibility and benefits may also apply.

Company

hirify.global designs, builds, manages, and modernizes mission-critical technology systems for enterprise customers.

What you will do

  • Design, optimize, and maintain scalable ETL/ELT pipelines for high-volume batch processing and low-latency streaming.
  • Architect distributed systems, database models, and cloud-native resources across hybrid environments.
  • Deploy vector databases, semantic layers, RAG search capabilities, agent memory structures, and feature stores for AI and machine learning workloads.
  • Build secure, high-throughput APIs that connect autonomous systems with distributed datasets.
  • Implement data quality, observability, metadata lineage, automated testing, and schema-drift prevention.
  • Capture deployment insights and collaborate with global teams to evolve platforms while reducing technical debt.

Requirements

  • Strong production-grade Python coding skills and advanced SQL optimization experience.
  • Experience designing and orchestrating ETL/ELT pipelines with tools such as Airflow, dbt, and Kafka.
  • Expertise in database modeling, distributed computing, and cloud-native environments including AWS, Azure, or GCP.
  • Hands-on experience with vector databases such as Pinecone, Milvus, Chroma, or Weaviate and vector indexing strategies.
  • Knowledge of data quality, pipeline lineage, and observability tools such as Great Expectations and DataHub.
  • Bachelor’s degree in a relevant technical field or equivalent professional experience; cloud certification or willingness to obtain one is required.

Nice to have

  • Experience with regulated industries such as financial services, healthcare, or the public sector.
  • Docker and Kubernetes experience.
  • Semantic modeling, metadata management, and AI agent integration experience.
  • CI/CD, Git, GitHub, and Agile delivery experience.
  • Advanced database, data architecture, or machine learning certifications, or a relevant master’s degree.

Culture & Benefits

  • Collaborative, customer-focused environment with a growth mindset and cross-functional delivery model.
  • Medical and dental coverage, disability and retirement benefits, paid leave, and paid time off.
  • Access to learning programs and professional certifications from providers including Microsoft, Google, Amazon, and Skillsoft.
  • Inclusive workplace supported by employee inclusion networks and company-wide volunteering programs.
  • COVID-19 vaccination is not generally required, although client or partner sites may have their own vaccination policies.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →