Назад
Company hidden
3 дня назад

Software Engineer 3 (Enterprise AI)

151 000 - 227 000CAD
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Singapore/US/Serbia +7 еще
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Software Engineer 3 (Enterprise AI) (Kubernetes/LLM): Building horizontally scalable, fault-tolerant platform and inference services for enterprise AI workloads with an accent on distributed systems, multi-tenant architectures, and low-latency LLM serving. Focus on designing request routing, token streaming, observability, CI/CD automation, and reliable deployments across on-premises, hybrid, and cloud environments.

Location: Hybrid, with at least 3 days per week onsite in applicable locations including Vancouver, San Jose, Durham, Mexico City, Bangalore, Pune, Hoofddorp, Belgrade, Barcelona, Singapore, Sydney, and Tokyo.

Salary: CAD $151,000–$227,000 per year, plus potential equity, bonuses, and benefits.

Company

hirify.global develops the hirify.global Cloud Platform for enterprise AI, including infrastructure for generative AI, large language models, and agentic AI applications.

What you will do

  • Architect and develop scalable, containerized, fault-tolerant services on Kubernetes for enterprise AI and LLM workloads.
  • Build high-performance inference and platform services supporting low-latency, high-throughput generative and agentic AI applications.
  • Design distributed systems across compute, storage, networking, virtualization, and multi-tenant infrastructure.
  • Implement observability architectures using Prometheus, Grafana, Datadog, OpenTelemetry, and related tools.
  • Develop CI/CD pipelines, deployment automation, and foundational LLM serving capabilities such as routing, rate limiting, token streaming, load balancing, and quota management.
  • Diagnose production issues, improve reliability, review designs and code, and collaborate with product, AI, and engineering teams.

Requirements

  • 6+ years of experience developing maintainable, resilient software products in a product development organization.
  • Strong foundations in data structures, algorithms, operating systems, networking, distributed systems, and datacenter architecture.
  • Production backend development experience with Go, Python, C++, or Rust.
  • Hands-on experience with Docker, Kubernetes, cloud-native architectures, CI/CD, and deployments across on-premises, cloud, and hybrid environments.
  • Experience with performance-sensitive systems, distributed data stores, observability platforms, and multi-tenant services.
  • Familiarity with LLM serving concepts and modern LLM capabilities including reasoning workflows, tool calling, prompt templates, and agents; a master's degree in computer science or equivalent practical experience.

Nice to have

  • Experience with PyTorch, TensorFlow, GPU acceleration, vLLM, DeepSpeed, Hugging Face TGI, or Triton.
  • Experience with RAG, vector databases, AI orchestration frameworks, production LLM APIs, agentic systems, or inference infrastructure.
  • Open-source contributions or experience working in large distributed codebases.

Culture & Benefits

  • Hybrid work with at least three onsite days per week where the workplace policy applies.
  • RRSP with dollar-for-dollar matching up to 7% of base salary in Vancouver.
  • Mental health coverage and paramedical benefits.
  • Paid maternity and parental leave, bereavement leave, RSUs, and an employee stock purchase plan with a 15% discount.
  • Collaboration with globally distributed product management, AI, and engineering teams.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →