Qualtiy Assurance Engineer- AI Products

Xpansiv Limited

Sheffield

Hybrid

GBP 60,000 - 70,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Xpansiv Limited is seeking a QA AI Engineer to test AI-enabled products and the proprietary LLM infrastructure, ensuring quality, reliability, and safety across systems and applications.

The role involves creating test plans, executing test cases, building evaluation pipelines, and validating both deterministic software behavior and non-deterministic AI outputs within an AI governance framework.

Qualifications

  • 10+ years of experience in software quality assurance, including manual and automated testing.

Responsibilities

  • Create test plans for AI-powered and traditional software, including manual, automated, performance, regression, security, and end-to-end testing.

Skills

Quality assurance
AI testing
Automation testing
Regression testing
Cross-team collaboration
Scripting
Git
Cypress
Playwright
Selenium
Rest Assured
Postman

Tools

Cypress
Playwright
Selenium
Rest Assured
Postman

Job description

Xpansiv is the leading infrastructure provider for the energy transition markets.

Our comprehensive platform includes registries, online marketplaces, market execution services, wholesale power solutions, and market data for energy and environmental commodity markets. Trusted worldwide, Xpansiv enables stakeholders to deliver transparent, credible, and auditable environmental claims to address the growing global demand for assurance and accountability on climate action and sustainability performance.

From our founding in 2009 through more than 10 acquisitions, Xpansiv has become a global leader in environmental commodity markets. We are backed by Blackstone and other leading investors.

Position Summary:

Seeking a QA AI Engineer to perform testing and quality assurance for systems, applications, and AI-enabled products developed by Xpansiv. This role is responsible for ensuring the quality, reliability, accuracy, and safety of Xpansiv’s AI-driven products and proprietary LLM infrastructure. The QA AI Engineer will create test plans, document and execute test cases, build automated and semi-automated evaluation frameworks, and validate both deterministic software behavior and non-deterministic AI outputs. This role works closely with AI Engineering, Product, Operations, Engineering, business analysts, business owners, and subject matter experts to build quality into AI products from the start and provide the human-in-the-loop assurance required by Xpansiv’s AI governance standards.

Key Responsibilities:
  • Create test plans for AI-powered and traditional software applications, including manual testing, automated testing, performance testing, regression testing, security testing, and end-to-end testing
  • Formulate and document test cases based on product requirements, user stories, acceptance criteria, AI governance standards, and business workflow expectations
  • Execute test cases through targeted manual testing, automated testing, exploratory testing, and AI-specific evaluation methods
  • Design, build, and maintain evaluation pipelines for AI-powered applications across Xpansiv business lines
  • Develop evaluation datasets, golden sets, and scenario suites that measure accuracy, consistency, structured-output quality, policy adherence, and business-rule compliance of LLM outputs
  • Detect, document, and reproduce AI-specific failure modes, including hallucinations, prompt injection, inconsistent outputs, formatting errors, data leakage, bias, unsafe responses, and model or prompt-update regressions
  • Build automated and semi-automated AI evaluation frameworks, including model-graded assertions, regression harnesses, prompt test suites, and quality scorecards, alongside traditional QA automation
  • Validate microservices, APIs, data extraction workflows, document-processing pipelines, RAG-based systems, agents, and structured-output generation for business-critical use cases
  • Perform functional, end-to-end, cross-browser, regression, security, performance, API, and integration testing as needed for AI-enabled and non-AI system components
  • Own quality gates and go/no-go readiness criteria for pilots, beta launches, production go-lives, and post-release model or prompt updates
  • Establish and track quality KPIs, including test coverage, pass rates, defect density, escaped-defect rate, AI accuracy metrics, hallucination rate, evaluation-score trends, and release readiness
  • Partner with AI Engineering to embed testing, monitoring, observability, evaluation, and quality controls into the AI development lifecycle
  • Conduct safety, bias, adversarial, and red-team testing to support responsible and compliant AI behavior aligned to Xpansiv’s AI Usage Standard
  • Work with developers, product managers, business analysts, business owners, and subject matter experts to review identified defects, provide clarifications, validate fixes, discuss solutions, and continuously improve AI accuracy and reliability
  • Support human-in-the-loop review requirements for high-risk client, financial, regulatory, and operational workflows
Qualifications:
  • 10+ years of experience in software quality assurance, including both manual and automated testing
  • Experience creating test plans, documenting test cases, executing test cases, validating defects, and supporting release-readiness decisions across multiple projects
  • Experience with functional testing, end-to-end testing, cross-browser testing, regression testing, security testing, performance testing, API testing, and integration testing
  • Experience testing or evaluating AI, ML, or LLM-powered systems, such as chatbots, copilots, agents, classifiers, RAG applications, document-processing workflows, or decision-support tools
  • Familiarity with generative AI evaluation methods, including golden datasets, model-graded evaluations, prompt regression testing, hallucination detection, red-teaming, and quality-drift monitoring
  • Working knowledge of prompt engineering, Retrieval Augmented Generation, agent-based systems, structured outputs, and AI guardrails sufficient to test these systems effectively
  • Experience with automated software testing tools such as Cypress, Playwright, Selenium, Rest Assured, or similar frameworks
  • Experience testing microservices, APIs, and web services using tools such as Postman, SoapUI, Rest Assured, or similar API testing tools
  • Proficiency in at least one scripting or programming language used for test automation, data validation, or evaluation tooling
  • Experience with source version control tools such as Git
  • Experience working in agile software development process models
  • Experience interacting with developers, product managers, business analysts, business owners, and subject matter experts to review and validate test plans, test cases, test results, and defects
  • Understanding of AI governance, responsible AI principles, data-handling guardrails, privacy considerations, and human-in-the-loop review practices
Skills / Abilities:
  • Strong communication and collaboration skills with clients, developers, business analysts, product managers, business owners, management, and cross-functional stakeholders
  • Strong analytical, troubleshooting, and failure-mode thinking, with the ability to translate ambiguous requirements and AI behavior into concrete test scenarios
  • Exceptional attention to detail, time management, and documentation quality
  • Proven ability to work independently as a self-starter and collaboratively as part of a cross-functional team
  • Ability to quickly absorb business and technical concepts, including unfamiliar AI workflows, data flows, and domain-specific business rules
  • Proven ability to work calmly under tight deadlines, production-readiness pressure, or critical quality situations
  • Curiosity and sound judgment when evaluating emerging AI capabilities, limitations, risks, and quality tradeoffs
What can you expect throughout the interview process:
  • Step 1- Take-home assessment
  • Step 2- Recruiter screening
  • Step 3- Technical Interview with the team
  • Step 4- Key Stakeholder + Final hiring manager interview
Base Salary

Compensation for this role will vary among specific regions due to geographic differentials in the labor market, actual pay will be determined considering factors such as relevant skills and experience, knowledge, education and training. However, compensation range for this role is expected to be as follows:

£60,000-£70,000

Here at Xpansiv, we cultivate diversity, celebrate individuality, and believe unique perspectives are key to our collective success in building trust and transparency in global efforts toward net-zero future. Xpansiv is committed to equal employment opportunity regardless of race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, disability, genetic information, protected veteran status, or any status protected by applicable federal, state, or local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Quality Assurance Engineer- AI Products
Quality Assurance Engineer- AI Products

Xpansiv • Sheffield

On-site
GBP 60,000 - 70,000
AI Quality Engineer
AI Quality Engineer

Xapien • Greater London

On-site
GBP 90,000 - 130,000
Private health insurance
Life Insurance
Unlimited holidays
+2
AI Quality Engineer
AI Quality Engineer

Xapien • City Of London

On-site
NOK 895,000 - 1,279,000
Private health insurance
Life Insurance
Unlimited holidays
+2
QA AI Engineer for Responsible AI and ML Systems
QA AI Engineer for Responsible AI and ML Systems

Xpansiv Limited • Sheffield

Hybrid
GBP 60,000 - 70,000
AI Quality Engineer
AI Quality Engineer

SoCode Recruitment • Greater London

Hybrid
GBP 70,000 - 90,000
Employee shares
Private health insurance
Unlimited holidays
+1
Quality Engineer
Quality Engineer

Xplor Education • Newcastle upon Tyne

On-site
GBP 45,000 - 75,000
Paid Parental Leave
#GiveBackDays/Commitment to social/vol
Diversity & Inclusion initiatives
+2
AI-first QA Engineer
AI-first QA Engineer

United States Digital Space LLC • United Kingdom

On-site
GBP 65,000 - 90,000
Remote-friendly work environment
Global distributed team
AI-focused organization
+3
Quality Engineer
Quality Engineer

Xplor • Newcastle upon Tyne

On-site
GBP 55,000 - 85,000
Paid Parental Leave
#GiveBackDays – 3 days volunteering
Diversity & Inclusion initiatives
+2
Principal AI Quality Engineer
Principal AI Quality Engineer

Fourth • Greater London

Hybrid
GBP 120,000 - 180,000
Holiday allowance
Hybrid work
Gym discounts
+8
QA AI Engineer for Trustworthy AI & Platform Quality
QA AI Engineer for Trustworthy AI & Platform Quality

Xpansiv • Sheffield

On-site
GBP 60,000 - 70,000