Назад
Company hidden
5 дней назад

Data Engineer, Forward Deployed Engineer (AI)

143 640 - 273 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Data Engineer, Forward Deployed Engineer (AI): Building and deploying scalable data infrastructure, ETL/ELT pipelines, vector databases, and context engines for customer-facing AI and machine learning systems with an accent on real-time streaming, retrieval-augmented generation, and production reliability. Focus on designing distributed data architectures, optimizing performance, implementing observability and compliance controls, and converting field prototypes into hardened enterprise platforms.

Location: Hybrid role requiring at least 3 days per week on-site in Frisco, Texas; candidates must reside within a commutable distance. Relocation assistance is not provided.

Salary: $143,640–$273,000 annually in the United States, with location-specific ranges for certain states and cities.

Company

hirify.global designs, builds, manages, and modernizes mission-critical technology systems for enterprise customers.

What you will do

  • Design, optimize, and maintain scalable ETL/ELT pipelines for high-volume batch processing and low-latency streaming.
  • Architect distributed systems, database models, and cloud-native resources across hybrid and public-cloud environments.
  • Deploy vector databases, semantic layers, RAG search systems, agent memory structures, APIs, and feature stores for AI and machine learning workloads.
  • Implement data quality, metadata lineage, automated testing, observability, and schema-drift prevention.
  • Diagnose performance bottlenecks and ensure solutions meet security, compliance, and data-protection requirements.
  • Capture deployment insights, collaborate with global practices, and evolve prototypes into stable customer platforms.

Requirements

  • Production-grade Python coding skills and advanced SQL optimization experience.
  • Experience building and orchestrating ETL/ELT pipelines with tools such as Airflow, dbt, and Kafka.
  • Expertise in database modeling, distributed computing, and AWS, Azure, or GCP cloud-native deployments.
  • Hands-on experience with vector databases such as Pinecone, Milvus, Chroma, or Weaviate and vector indexing strategies.
  • Knowledge of data quality, pipeline lineage, and observability tools such as Great Expectations and DataHub.
  • Bachelor’s degree in a relevant technical field or equivalent professional experience; cloud certifications or willingness to obtain them is required.

Nice to have

  • Experience with regulated industries such as financial services, healthcare, or the public sector.
  • Docker and Kubernetes deployment experience.
  • Semantic modeling, metadata management, and AI agent architecture experience.
  • CI/CD, Git, GitHub, and Agile delivery experience.
  • Advanced database, data architecture, or machine learning certifications, or a related master’s degree.

Culture & Benefits

  • Collaborative, customer-focused environment with a growth mindset and hands-on delivery model.
  • Medical and dental coverage, disability and retirement benefits, paid leave, and paid time off.
  • Discretionary annual bonus program based on performance.
  • Learning programs and access to certification opportunities from Microsoft, Google, Amazon, Skillsoft, and others.
  • Inclusive workplace supported by employee inclusion networks and volunteering programs.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →