Назад
Company hidden
2 дня назад

Qualtiy Assurance Engineer- AI Products

Формат работы
remote (только United_kingdom)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Qualtiy Assurance Engineer- AI Products (AI/QA): Testing and assuring the quality, reliability, accuracy, and safety of AI-enabled products and LLM infrastructure with an accent on automated evaluation, traditional QA automation, and responsible AI testing. Focus on building evaluation datasets and regression frameworks, detecting hallucinations and prompt-injection failures, and validating APIs, RAG systems, agents, and structured outputs.

Location: Remote, Sheffield, United Kingdom

Company

Develops AI-enabled products and proprietary LLM infrastructure requiring software quality assurance and AI governance controls.

What you will do

  • Create and execute manual, automated, exploratory, functional, regression, performance, security, API, integration, and end-to-end tests.
  • Design and maintain AI evaluation pipelines, datasets, golden sets, scenario suites, regression harnesses, prompt test suites, and quality scorecards.
  • Validate microservices, APIs, data extraction workflows, document-processing pipelines, RAG systems, agents, and structured-output generation.
  • Detect, reproduce, and document AI failure modes including hallucinations, prompt injection, inconsistent outputs, data leakage, bias, unsafe responses, and regressions.
  • Own quality gates and go/no-go readiness criteria for pilots, launches, production releases, and model or prompt updates.
  • Partner with engineering, product, business, and subject-matter stakeholders on defects, fixes, monitoring, observability, and human-in-the-loop assurance.

Requirements

  • 10+ years of software quality assurance experience across manual and automated testing.
  • Experience creating test plans and cases, validating defects, measuring quality KPIs, and supporting release-readiness decisions.
  • Experience testing AI, ML, or LLM-powered systems and using generative AI evaluation methods such as golden datasets, model-graded evaluations, prompt regression testing, hallucination detection, and red-teaming.
  • Working knowledge of prompt engineering, Retrieval Augmented Generation, agent-based systems, structured outputs, AI guardrails, responsible AI, privacy, and human-in-the-loop practices.
  • Experience with Cypress, Playwright, Selenium, Rest Assured, Postman, SoapUI, or similar tools; proficiency in at least one scripting or programming language and Git.
  • Experience working in agile environments and collaborating with developers, product managers, analysts, business owners, and subject-matter experts.

Culture & Benefits

  • Remote work arrangement.
  • Cross-functional collaboration across AI engineering, product, operations, engineering, business analysis, and subject-matter expertise.
  • Focus on responsible, compliant, reliable, and safe AI behavior.
  • Base salary is referenced, but no amount is provided.

Hiring process

  • The interview process includes evaluation of the candidate’s QA, AI testing, automation, and release-readiness experience.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →