Назад
Company hidden
обновлено 22 часа назад

Platform Power Thermal Performance Engineer (AI Infrastructure)

Формат работы
onsite
Тип работы
fulltime
Грейд
trainee
Английский
b2
Страна
China
Вакансия из списка Hirify.GlobalВакансия из Hirify RU Global, списка компаний с восточно-европейскими корнями
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Platform Power Thermal Performance Engineer (AI Infrastructure): Validating and optimizing performance for AI server platforms across CPU hosts, GPU I/O, memory, storage, and networking with an accent on benchmark execution, topology validation, and system-level analysis. Focus on debugging performance bottlenecks, automating experiments, and supporting scale-up and scale-out platform optimization.

Location: On-site in Shanghai, China

Company

hirify.global develops computing technologies and AI server platforms.

What you will do

  • Validate topologies and analyze performance across CPU hosts, GPUs, memory, storage, and network components.
  • Develop validation plans, execute benchmarks, collect data, and analyze results for CPU host and GPU I/O studies.
  • Investigate performance issues with senior engineers and propose improvements from the host platform perspective.
  • Collaborate with cross-functional teams to communicate test progress, document technical findings, and support issue closure.

Requirements

  • Upcoming graduate with a Master of Science degree or higher in Electrical Engineering, Computer Science, Computer Engineering, or a related technical field.
  • Fundamentals in computer architecture, operating systems, or computer systems, including interest in CPU, memory, and I/O behavior.
  • Basic knowledge of GPU computing, parallel programming, or accelerator-based systems.
  • Interest in AI server and cluster-level architecture, networking, storage, scale-up, and scale-out design.
  • Programming and scripting skills in Python, C/C++, or shell scripting for test automation, data processing, and performance analysis.
  • Good verbal and written English communication skills are required.

Nice to have

  • Exposure to PCIe/CXL, coherency, IOMMU, NUMA, or host-device data movement.
  • Experience with CUDA, ROCm, OpenCL, SYCL, or related GPU and AI accelerator frameworks.
  • Experience with system validation, performance testing, Linux development, machine learning workloads, networking, storage, or GPU computing through internships, research, coursework, or academic projects.

Culture & Benefits

  • Hands-on learning opportunities in modern AI infrastructure and server performance engineering.
  • Work with experienced engineers and cross-functional technical teams.
  • Exposure to single-server performance tuning and cluster-scale design concepts.
  • Full-time college graduate position with on-site work in Shanghai.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →