Back
2 months ago

Staff Engineer, Datacenter Server Lifecycle (AI)

Work format
hybrid
Work type
fulltime
Grade
senior
English
b2
Country
Australia
This vacancy is from Hirify.Global listVacancy from Hirify Global, list of international tech companies
Plus is required to make matches and apply

Match & Cover letter

Plus required for matching with this vacancy

Job description

Text:
/
TL;DR
Staff Engineer, Datacenter Server Lifecycle (AI): Designing and owning the end-to-end operational journey of servers in global facilities with an accent on provisioning, deployment, and hardware security. Focus on building automation for tens of thousands of servers, ensuring verified chains of integrity, and innovating AI-native datacenter workflows.

Location: Sydney, Australia (Hybrid - minimum 25% office attendance)

Company

Anthropic is a public benefit corporation dedicated to creating reliable, interpretable, and steerable AI systems.

What you will do

  • Build automation to support and manage datacenters containing tens of thousands of servers.
  • Define and own the end-to-end system lifecycle strategy, from provisioning and deployment to decommissioning.
  • Partner with Infrastructure Security to design and enforce trusted compute standards across the server lifecycle.
  • Collaborate with the Networking team to ensure seamless end-to-end connectivity across all sites.
  • Develop and maintain tooling to track machine health, configuration, and operational status across the fleet.

Requirements

  • Hands-on experience with server hardware, rack deployment, cabling, and troubleshooting at scale.
  • Deep understanding of hardware lifecycle management, including asset tracking and provisioning workflows.
  • Proficiency in at least one programming language such as Python, Rust, Go, or Java.
  • Working knowledge of Kubernetes and large-scale cloud providers (AWS, Azure, GCP).
  • Must be based in Sydney, Australia to comply with the hybrid office policy.
  • Willingness to travel occasionally to datacenter sites across North America.

Nice to have

  • 8+ years of experience in datacenter infrastructure management or related disciplines.
  • Experience with GPU or AI accelerator hardware (e.g., NVIDIA A100/H100, Google TPUs).
  • Familiarity with LinuxBoot, NixOS, or similar modern provisioning and OS tooling.
  • Background in large-scale capacity planning and hardware refresh at a hyperscaler.
  • Expertise in hardware security concepts like Secure Boot, TPM, and hardware attestation.

Culture & Benefits

  • Competitive compensation and benefits with optional equity donation matching.
  • Generous vacation and parental leave.
  • Flexible working hours and high-quality collaborative office spaces.
  • Visa sponsorship available for qualified candidates.
  • Strongly collaborative environment focused on high-impact AI research.

Be careful: if the employer asks you to log into their system using iCloud/Google, send codes/passwords, or run code/software, don't do it - these are scammers. Always click "Report" or contact support. More in guide →