Назад
5 дней назад

Staff Software Engineer (Platform Metering)

18 333 - 22 222$
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Staff Software Engineer (Platform Metering): Building the source of truth for GPU cloud resource consumption to power usage-based billing and credit enforcement with an accent on distributed systems and event-driven architecture. Focus on designing accurate metering models, ensuring idempotency, and creating scalable standards for resource tracking.

Location: Must be based in Houston, New York, San Francisco, or Seattle

Salary: $220,000 - $266,667 USD

Company

Nscale provides a GPU cloud platform for compute and storage resources tailored for AI and ML workloads.

What you will do

  • Define domain-level technical direction for platform metering across GPU compute and storage.
  • Design and implement reliable event producers and resource watchers for Kubernetes, Slurm, and bare metal.
  • Ensure system idempotency, auditability, and reconciliation to prevent billing disputes and data loss.
  • Integrate usage data with billing systems for rating, credit burn-down, and cost attribution.
  • Establish shared schemas and operational standards so new platform services can implement metering consistently.
  • Mentor other engineers through architecture reviews and implementation of high-integrity usage systems.

Requirements

  • Extensive experience designing and operating production distributed systems in cloud infrastructure, billing, or platform engineering.
  • Strong knowledge of resource lifecycle and usage tracking (Kubernetes nodes, compute instances, storage volumes).
  • Proficiency in typed backend languages, primarily Go, with some Rust or Python.
  • Deep understanding of event-driven architecture, including reliable delivery, idempotency, and failure handling.
  • Proven ability to lead ambiguous technical work across team boundaries through influence.
  • Must be based in Houston, New York, San Francisco, or Seattle.

Nice to have

  • Experience with GPU cloud infrastructure, workload scheduling, or AI/ML training workloads.
  • Background in ledger-style, event-sourced, or replayable data systems.
  • Experience with Kubernetes controllers, operators, CRDs, and admission webhooks.
  • Product sense for translating complex usage data into intuitive dashboards for customers and finance.

Culture & Benefits

  • Competitive benefits package including medical, dental, and vision insurance.
  • Flexible paid time off and parental leave.
  • Retirement plan participation.
  • High-ownership environment with a bias toward measurable business impact.
  • Modern development workflow utilizing AI tools like Claude and Cursor to increase leverage.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →