Agentic QA Engineer

Qentelli LLC

Hyderabad

On-site

INR 1,800,000 - 3,000,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Qentelli LLC in Hyderabad, India seeks a hands-on AI Engineer to design and execute end-to-end testing for agentic AI solutions and multi-agent systems in production-grade environments. You will partner with the Agentic Operations Team to ensure resiliency, reliability, accuracy, latency, orchestration correctness, and scale, from development to production.

You will establish QA frameworks, build reusable test artifacts, drive macro-level validations across complex workflows, mentor QA

Qualifications

  • 5+ years in Software QA/Testing, with 2+ years in AI/ML or LLM-based systems.
  • Hands-on testing experience for agentic/multi-agent architectures.
  • Strong Python programming for test harnesses, simulators, and fixtures.
  • Experience with LLM evaluation metrics and guardrails.
  • Expertise in distributed systems testing, latency profiling, and chaos engineering.
  • Familiarity with orchestration frameworks and cloud-native tools.

Responsibilities

  • Define and own the QA strategy for agentic/multi-agent AI systems across dev, staging, and prod.
  • Mentor a team of QA engineers; establish testing standards and review practices.
  • Partner with Agentic Operations, Data Science, MLOps, and Platform teams to embed QA in the SDLC and incident response.
  • Design tests for agent orchestration, tool calling, planner-executor loops, and inter-agent coordination.
  • Validate state management, context windows, memory stores, and prompt/graph correctness.
  • Implement scenario fuzzing and resilience testing including chaos experiments and controlled outages.
  • Create macro validation frameworks for complex multi-step agent workflows.
  • Instrument guardrail validations (toxicity, PII, hallucination, policy compliance).

Skills

Python programming
LLM evaluation
Distributed systems testing
Chaos engineering
Cross-functional leadership
Communications

Tools

GitHub Actions
Azure DevOps
OpenTelemetry
Prometheus
Grafana
Datadog
LangChain
LangGraph
LlamaIndex
DSPy

Job description

We are seeking a hands-on AI Engineer to design and execute end-to-end testing strategies for agentic AI solutions, including multi-agent systems in production-grade environments. This role partners with the Agentic Operations Team to ensure resiliency, reliability, accuracy, latency, orchestration correctness, and scale. You will establish QA frameworks, build reusable test artifacts, drive macro-level validations across complex workflows, and lead the QA function for Agentic AI from Dev to Prod.

Key Responsibilities
  • Agentic & MultiAgent Testing
  • Reliability, Resiliency, and Latency
  • Accuracy & Macro-Level Validations
  • Scale & Orchestration
  • Dev Prod Readiness
  • Define and own the QA strategy for agentic/multi-agent AI systems across dev, staging, and prod.
  • Mentor a team of QA engineers; establish testing standards, coding guidelines for test harnesses, and review practices.
  • Partner with Agentic Operations, Data Science, MLOps, and Platform teams to embed QA in the SDLC and incident response.
  • Design tests for agent orchestration, tool calling, planner-executor loops, and inter-agent coordination (e.g., task decomposition, handoff integrity, and convergence to goals).
  • Validate state management, context windows, memory/knowledge stores, and prompt/graph correctness under varying conditions.
  • Implement scenario fuzzing (e.g., adversarial inputs, prompt perturbations, tool latency spikes, degraded APIs).
  • Create resilience testing suites: chaos experiments, failover, retries/backoff, circuit-breaking, and degraded mode behavior.
  • Establish latency SLOs and measure end-to-end response times across orchestration layers (LLM calls, tool invocations, queues).
  • Ensure reliability through soak tests, canary verifications, and automated rollbacks.
  • Define ground-truth and reference pipelines for task accuracy (exact match, semantic similarity, factuality checks).
  • Build macro validation frameworks that validate task outcomes across multi-step agent workflows (e.g., complex data pipelines, content generation + verification agent loops).
  • Instrument guardrail validations (toxicity, PII, hallucination, policy compliance).
  • Design load/stress tests for multi-agent graphs under scale (concurrency, throughput, queue depth, backpressure).
  • Validate orchestrator correctness (DAG execution, retries, branching, timeouts, compensation paths).
  • Engineer reusable test artifacts (scenario configs, synthetic datasets, prompt libraries, agent graph fixtures, simulators).
  • Integrate tests into CI/CD (pre-merge gates, nightly, canary) and production monitoring with alerting tied to KPIs.
  • Define release criteria and run operational readiness (performance, security, compliance, cost/latency budgets).
Required Qualifications
  • 5+ years in Software QA/Testing, with 2+ years in AI/ML or LLM-based systems; hands-on experience testing agentic/multi-agent architectures.
  • Strong programming skills in Python experience building test harnesses, simulators, and fixtures.
  • Experience with LLM evaluation (exact/soft match, BLEU/ROUGE, BERTScore, semantic similarity via embeddings), guardrails, and prompt testing.
  • Expertise in distributed systems testing latency profiling, resiliency patterns (circuit breakers, retries), chaos engineering, and message queues.
  • Familiarity with orchestration frameworks (LangChain, LangGraph, LlamaIndex, DSPy, OpenAI Assistants/Actions, Azure OpenAI orchestration, or similar).
  • Proficiency with CI/CD (GitHub Actions/Azure DevOps), observability (OpenTelemetry, Prometheus/Grafana, Datadog), and feature flags/canaries.
  • Solid understanding of privacy/security/compliance in AI systems (PII handling, content policies, model safety).
  • Excellent communication and leadership skills; proven ability to work cross-functionally with Ops, Data, and Engineering.
Preferred Qualifications
  • Experience with multi-agent simulators, agent graph testing, and tooling latency emulation.
  • Knowledge of MLOps (model versioning, datasets, evaluation pipelines) and A/B experimentation for LLMs.
  • Background in cloud (AWS), serverless, containerization, and event-driven architectures.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Engineer
AI Engineer

Qentelli • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Agentic AI Test Engineer
Agentic AI Test Engineer

Deutsche Telekom Digital Labs • Gurugram District

On-site
INR 1,620,000 - 1,980,000
Quality Intelligence Engineer _AI-Powered Quality Engineering & Agent
Quality Intelligence Engineer _AI-Powered Quality Engineering & Agent

Capgemini • Hyderabad

Hybrid
INR 1,400,000 - 2,200,000
Principal Software Engineer
Principal Software Engineer

Cadence • Bengaluru

On-site
INR 4,500,000 - 7,500,000
Software Development Engineer - Test
Software Development Engineer - Test

sghcorp.com • Bengaluru

On-site
INR 11,561,000 - 17,341,000
AI/LLM QA Engineer - Agentic and Multi-Agent System Testing
AI/LLM QA Engineer - Agentic and Multi-Agent System Testing

Crew Kraftorz LLP • Hyderabad

On-site
INR 1,500,000 - 2,200,000
Sr. Software Test Automation Engineer - Playwright with TypeScript and Agentic AI
Sr. Software Test Automation Engineer - Playwright with TypeScript and Agentic AI

wk • India

On-site
INR 1,000,000 - 1,800,000
QA Engineer – Automation & Agentic Applications
QA Engineer – Automation & Agentic Applications

Beroe Inc • India

On-site
INR 1,000,000 - 1,500,000
Lead QA - Agentic AI
Lead QA - Agentic AI

Sycamore Informatics Inc. • India

On-site
INR 3,500,000 - 6,000,000
Sr AI/ML Engineer
Sr AI/ML Engineer

Staples India • Chennai District

On-site
INR 3,200,000 - 5,200,000