Назад
Company hidden
2 дня назад

Senior QA Engineer (Performance & Reliability)

100 000 - 160 000$
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior QA Engineer – Performance & Reliability (Embedded Systems/AI Security): Designing and executing performance, stress, endurance, thermal, and power validation for Secure TCU and BMC systems with an accent on bottleneck analysis, reliability testing, and system-level debugging. Focus on investigating failures with firmware and hardware teams, building automated test frameworks, and analyzing complex performance and reliability issues.

Location: San Jose, California, United States

Salary: USD 100,000–160,000 per year

Company

hirify.global develops Trusted Control/Compute Unit solutions that combine hardware security, AI-driven resilience, efficiency, and real-time platform management for data centers and cloud infrastructure.

What you will do

  • Design and execute performance benchmarks, stress tests, longevity and endurance tests, and thermal and power characterization for TCU and BMC systems.
  • Analyze heavy workloads to identify throughput, latency, and CPU, memory, and PCIe utilization bottlenecks.
  • Conduct MTBF prediction, long-duration stability testing, and error-injection campaigns.
  • Lead root-cause investigations using oscilloscopes, logic analyzers, firmware traces, crash dumps, and other debugging tools.
  • Collaborate with firmware and hardware engineers to reproduce issues and verify fixes.
  • Develop automated performance testing frameworks, reporting dashboards, and reliability metrics.

Requirements

  • 5+ years of experience in embedded system testing, performance verification, and reliability engineering.
  • Knowledge of TCU, BMC, HMC, Root of Trust, Secure Boot, TPM, HSM, PCIe Gen4/5, DDR memory, and networking protocols.
  • Experience with performance profiling tools, traffic generators, standard benchmarks, and thermal and power measurement tools.
  • Strong Python scripting skills for test automation and data analysis; familiarity with C/C++ for code analysis.
  • Strong Linux/Unix skills, including kernel tuning, system monitoring, and log analysis.
  • Experience with Jenkins, GitLab CI, Git, and a BS/MS degree in Computer Science, Electrical Engineering, or a related field.

Nice to have

  • Experience with AI-driven log analysis or anomaly detection for reliability prediction.
  • Validation experience with PCIe, CXL, DDR5, or LPDDR5.
  • Background in data center server architecture and thermal management.
  • Knowledge of Telcordia or JEDEC reliability standards.

Culture & Benefits

  • Work in a fast-growing semiconductor, systems, and solutions company.
  • Collaborate across hardware, firmware, and systems engineering disciplines.
  • Focus on practical real-world security and reliability challenges.
  • Continuous learning, mutual support, curiosity, persistence, and problem-solving are valued.
  • Diverse and multifaceted work environment with equal employment opportunity.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →