Назад
Company hidden
17 часов назад

Software Engineer - Infrastructure (AI)

Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
India
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Software Engineer - Infrastructure (AI) (Kubernetes/Cloud): Building and operating distributed infrastructure for AI agent workloads and production applications with an accent on Kubernetes orchestration, cloud platforms, observability, and reliability. Focus on developing Kubernetes controllers, automating infrastructure lifecycle operations, debugging complex cluster issues, and handling incident response across large-scale systems.

Location: Bangalore, India

Company

hirify.global builds autonomous coding agents that generate, test, and deploy production applications from plain-language intent. Its systems operate at global scale and support millions of applications.

What you will do

  • Maintain distributed microservices infrastructure running across Kubernetes and GCP/AWS.
  • Manage Kubernetes workloads with ArgoCD, including deployments, monitoring, troubleshooting, resource trees, and rollouts.
  • Operate CDN and edge infrastructure with Cloudflare for performance, caching, and traffic management.
  • Own observability with Grafana, Loki, Prometheus, and New Relic; improve monitoring, alerting, and distributed tracing.
  • Participate in PagerDuty on-call rotations, respond to incidents, and perform root cause analysis.
  • Support AI agent infrastructure, including scheduling, environment provisioning, deployments, trajectory tracking, and cost attribution; build internal tools and Kubernetes operators.

Requirements

  • 3+ years of software or platform engineering experience with production systems.
  • Production programming experience in Go or Python and hands-on experience building and deploying services on Kubernetes.
  • Experience with GitOps tooling such as ArgoCD or Flux.
  • Strong networking and DNS fundamentals, including TCP/IP, HTTP, load balancing, DNS resolution, TLS, and connectivity debugging.
  • Solid Linux/OS fundamentals and experience debugging with tools such as strace, tcpdump, and netstat.
  • Experience with relational and NoSQL databases, caching, message queues, streaming systems, cloud infrastructure, and at least one observability stack.

Nice to have

  • Experience building Kubernetes operators with kubebuilder, operator-sdk, or controller-runtime.
  • Experience tuning Kubernetes core components such as the API server, kubelet, or scheduler.
  • Familiarity with AI/LLM infrastructure, including token management, cost tracking, and agent orchestration.
  • Experience with CI/CD, infrastructure as code, large-scale distributed systems, or platform-as-a-service.
  • Startup experience and comfort working in fast-paced, ambiguous environments.

Culture & Benefits

  • Daily lunch and dinner provided.
  • Family insurance coverage of 3 lakhs.
  • Unlimited paid time off.
  • Flexible working hours.
  • Collaborative work with product and backend teams in a distributed, fast-moving environment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →