3 дня назад
Senior Cloud Support Engineer (AI Infrastructure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Cloud Support Engineer (AI Infrastructure): Supporting and troubleshooting Crusoe’s GPU-powered cloud infrastructure with an accent on virtual machines, hardware failures, scaling tests, and 24/7 customer support. Focus on diagnosing complex issues through CLI and internal tools, coordinating root cause analysis with SRE, networking, and storage teams, and improving support tooling and documentation.
Location: Tel Aviv, Israel; on-site
Company
builds vertically integrated AI infrastructure spanning energy, data centers, and cloud services for large-scale AI workloads.
What you will do
- Provide hands-on technical support through Zendesk while meeting SLAs and maintaining CSAT above 95%.
- Participate in a 24/7 support rotation and respond to critical infrastructure issues.
- Diagnose and resolve problems involving virtual machines, hardware failures, scaling tests, and node delivery testing.
- Manage alert triage and maintenance windows using CLI tools and internal systems.
- Collaborate with global SRE, networking, and storage teams from incident triage through root cause analysis.
- Create documentation, onboarding materials, and SOPs while improving support processes and tooling.
Requirements
- Bachelor’s degree in IT, computer science, engineering, or equivalent experience with 4+ years in a similar technical role.
- At least 5 years of experience in technical support, preferably in cloud, storage, or networking.
- Strong CLI skills and confidence working in Linux environments.
- Experience with Git, container orchestration such as Kubernetes, workload management tools such as Slurm and Terraform, and observability tools such as Grafana.
- Familiarity with AWS, Azure, or GCP and with high-performance computing technologies including InfiniBand, RDMA, RoCE, and SDN.
- Clear communication in English and Hebrew is required.
Nice to have
- Experience with NVMe, SSDs, distributed storage, block storage, object storage, or file storage.
- Familiarity with NFS, SMB, iSCSI, or NVMe-oF storage protocols.
- CKA, CKAD, CKS, KCNA, AWS, NVIDIA AI, or Linux Foundation certifications.
- Experience with automation tools or scripting languages.
- Interest in coaching teammates and improving technology efficiency and sustainability.
Culture & Benefits
- Mission-driven work focused on improving AI infrastructure through an energy-first approach.
- Close collaboration with engineering teams across energy, manufacturing, data center construction, and cloud services.
- High-ownership environment that values initiative, urgency, and creative problem-solving.
- Local-market benefits including pension contributions and additional work-life-balance perks.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Nebius
2 дня назад
Senior Support Engineer (AI)
5 дней назад
Technical Support Lead (AI)
4 дня назад
Customer Success Engineer Tel Aviv (Cybersecurity)
4 дня назад
Enterprise Customer Success Manager (SaaS)
5 дней назад
Technical Account Manager - Customer Success (Manufacturing)
4 дня назад
Customer Success Manager (SaaS)
75 000 - 95 000$