Lead II - Software Testing

TekWissen LLC

Bellevue (WA)

Hybrid

USD 120,000 - 160,000

Full time

12 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

TekWissen Group is seeking a Lead II - Software Testing in Bellevue, WA for a 6-month hybrid assignment. You will design and execute test plans for AI/ML features, build automated test suites, and evaluate model outputs for accuracy, bias, and robustness.

Collaboration with data scientists and ML engineers is essential to define acceptance criteria and quality metrics. The role requires strong Python, SQL, and experience with eval frameworks.

Qualifications

  • Experience designing and executing test plans for AI/ML features.
  • Strong Python scripting and data validation skills.
  • Familiarity with eval metrics for generative AI.

Responsibilities

  • Design and execute test plans for AI/ML-driven features, including model outputs, prompts, and integrated application behavior.
  • Build and maintain automated test suites covering functional, regression, integration, and API testing.
  • Evaluate model outputs for accuracy, consistency, bias, hallucination, and edge-case failures.
  • Develop evaluation frameworks and golden datasets/test cases to benchmark model performance over time.
  • Test prompt engineering changes, model version upgrades, and fine-tuning outputs for regressions.
  • Perform adversarial and red-team style testing to surface safety, security, and robustness issues.

Skills

AI/ML testing
Python
Test automation
SQL
CI/CD
Adversarial testing

Tools

Ragas
DeepEval
LangSmith
Promptfoo
OpenAI Evals
TruLens

Job description

Overview:

TekWissen is a global workforce management provider headquartered in Ann Arbor, Michigan that offers strategic talent solutions to our clients world-wide. Our client provider of digital technology and transformation, information technology and services

Position:

Lead II - Software Testing

Location:

Bellevue,WA 98008

Duration:

6 Months

Job Type:

Temporary Assignment

Work Type:

Hybrid

Job Description:
  • Design and execute test plans for AI/ML-driven features, including model outputs, prompts, and integrated application behavior
  • Build and maintain automated test suites covering functional, regression, integration, and API testing
  • Evaluate model outputs for accuracy, consistency, bias, hallucination, and edge-case failures
  • Develop evaluation frameworks and golden datasets/test cases to benchmark model performance over time
  • Test prompt engineering changes, model version upgrades, and fine-tuning outputs for regressions
  • Perform adversarial and red-team style testing to surface safety, security, and robustness issues
  • Validate data pipelines feeding into AI models (data quality, schema, drift detection)
  • Collaborate with data scientists/ML engineers to define acceptance criteria and quality metrics for models
  • Test latency, scalability, and reliability of AI services under load
  • Contribute to CI/CD pipelines, integrating automated and model-evaluation tests
  • Hands-on experience testing LLM-based products (chatbots, copilots, RAG systems, AI agents) - designing test cases for non-deterministic, generative outputs
  • Practical experience with AI/LLM evaluation frameworks (e.g., Ragas, DeepEval, LangSmith, Promptfoo, OpenAI Evals, TruLens) - building eval suites, scoring rubrics, and golden datasets
  • Working knowledge of eval metrics for generative AI: hallucination rate, faithfulness/groundedness, relevance, answer correctness, toxicity/bias scoring, BLEU/ROUGE/semantic similarity where applicable
  • Experience with prompt regression testing - validating prompt changes and model/version upgrades against baseline eval sets
  • Strong proficiency in Python for writing eval scripts, test harnesses, and data validation logic
  • Familiarity with SQL and data validation techniques

TekWissen Group is an equal opportunity employer supporting workforce diversity.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

QA Engineer (AI)
QA Engineer (AI)

Apptad Inc • Frisco (TX)

On-site
USD 120,000 - 180,000
QA Engineer for AI Products
QA Engineer for AI Products

Intone Inc • Bellevue (WA)

Hybrid
USD 120,000 - 160,000
QA Engineer – AI/ML
QA Engineer – AI/ML

Apptad Inc • Frisco (TX)

On-site
USD 120,000 - 160,000
AI/ML Testing Lead — QA for Generative Models
AI/ML Testing Lead — QA for Generative Models

TekWissen LLC • Bellevue (WA)

Hybrid
USD 120,000 - 160,000
SDET - Automation Engineer - AI
SDET - Automation Engineer - AI

Aegistech • New Jersey

Hybrid
USD 110,000 - 170,000
AI Agentic Tester
AI Agentic Tester

Delan Associates, Inc • St. Louis (MO)

On-site
USD 90,000 - 130,000
AI Software Engineer – LLM Evaluation & Automation (Remote)
AI Software Engineer – LLM Evaluation & Automation (Remote)

Stage 4 Solutions Inc • United States

Remote
USD 99,000 - 108,000
Health benefits
401K
Senior AI Test Automation Engineer
Senior AI Test Automation Engineer

Genuine Parts Company • Alabama

On-site
USD 120,000 - 150,000
AI Agentic Tester
AI Agentic Tester

Delan Associates, Inc • Denver (CO)

On-site
USD 90,000 - 120,000
AI Software Test Engineer
AI Software Test Engineer

Spectraforce Technologies • Ann Arbor (MI)

On-site
USD 90,000 - 130,000