Architect

WinWire

Bengaluru

On-site

INR 4,000,000 - 8,000,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

WinWire in Bengaluru, India seeks an experienced Architect to design end-to-end agentic systems, including task decomposition, tool design, memory models, and human-in-the-loop checkpoints. You will shape production RAG pipelines with chunking, hybrid retrieval, and grounding, while defining guardrails and cost-aware AWS deployments.

You will lead evaluation layers, establish traceable metrics, and collaborate with clients on discovery and value proofs. Strong Python and AWS experience required.

Qualifications

  • 2-3 year experience in Agentic system design
  • 2-3 year experience in Agent frameworks
  • 3 year experience in RAG engineering
  • AWS AI stack – 3 years of experience
  • 5 years of experience in AWS core services
  • 7+ years of experience in Python engineering
  • 2+ years of experience Evaluation and LLMOps
  • 2+ years of experience AI security and responsible AI

Responsibilities

  • Architect agentic solutions end to end: task decomposition, tool/function design, memory and state models, orchestration topology (supervisor, hierarchical, sequential, parallel), and human-in-the-loop checkpoints.
  • Design and tune production RAG pipelines — chunking, hybrid retrieval, reranking, query rewriting, metadata filtering, grounding and citation — and own the retrieval quality metrics, not just the pipeline diagram.
  • Build the evaluation layer alongside the system: golden datasets, trace-level assertions, LLM-as-judge rubrics, regression suites that gate deployment.
  • Select and justify the AWS service composition for each workload, with an explicit view on cost, latency, data residency, and failure modes.
  • Define guardrails and the security posture for agents that take real actions: least-privilege tool permissions, prompt injection defenses, PII handling, audit trails.
  • Set engineering standards for the practice — reference architectures, reusable agent and tool patterns, observability conventions — and mentor engineers into them.
  • Partner with client stakeholders on solution shaping, technical discovery, effort estimation, and proof-of-value scoping.

Skills

Agentic system design
Agent frameworks
RAG engineering
AWS AI stack
AWS core services
Python engineering
Evaluation & LLMOps
AI security

Tools

Strands Agents SDK
LangGraph
CrewAI
AutoGen
LlamaIndex

Job description

Architect agentic solutions end to end: task decomposition, tool/function design, memory and state models, orchestration topology (supervisor, hierarchical, sequential, parallel), and human-in-the-loop checkpoints. Design and tune production RAG pipelines — chunking, hybrid retrieval, reranking, query rewriting, metadata filtering, grounding and citation — and own the retrieval quality metrics, not just the pipeline diagram. Build the evaluation layer alongside the system: golden datasets, trace-level assertions, LLM-as-judge rubrics, regression suites that gate deployment. Select and justify the AWS service composition for each workload, with an explicit view on cost, latency, data residency, and failure modes. Define guardrails and the security posture for agents that take real actions: least-privilege tool permissions, prompt injection defenses, PII handling, audit trails. Set engineering standards for the practice — reference architectures, reusable agent and tool patterns, observability conventions — and mentor engineers into them. Partner with client stakeholders on solution shaping, technical discovery, effort estimation, and proof-of-value scoping.

Responsibilities
  • Architect agentic solutions end to end: task decomposition, tool/function design, memory and state models, orchestration topology (supervisor, hierarchical, sequential, parallel), and human-in-the-loop checkpoints.
  • Design and tune production RAG pipelines — chunking, hybrid retrieval, reranking, query rewriting, metadata filtering, grounding and citation — and own the retrieval quality metrics, not just the pipeline diagram.
  • Build the evaluation layer alongside the system: golden datasets, trace-level assertions, LLM-as-judge rubrics, regression suites that gate deployment.
  • Select and justify the AWS service composition for each workload, with an explicit view on cost, latency, data residency, and failure modes.
  • Define guardrails and the security posture for agents that take real actions: least-privilege tool permissions, prompt injection defenses, PII handling, audit trails.
  • Set engineering standards for the practice — reference architectures, reusable agent and tool patterns, observability conventions — and mentor engineers into them.
  • Partner with client stakeholders on solution shaping, technical discovery, effort estimation, and proof-of-value scoping.
Qualifications
  • 2-3 year experience in Agentic system design
  • 2-3 year experience in Agent frameworks
  • 3 year experience in RAG engineering
  • AWS AI stack – 3 years of experience
  • 5 years of experience in AWS core services
  • 7+ years of experience in Python engineering
  • 2+ years of experience Evaluation and LLMOps
  • 2+ years of experience AI security and responsible AI
Required Skills
  • Agentic system design: Planning and execution loops, ReAct and plan-execute patterns, tool/function calling design, short- and long-term memory, state persistence and recovery, multi-agent orchestration and handoff design, MCP and A2A for tool and agent interoperability. Practical judgment on when an agent is the wrong answer and a deterministic workflow is the right one.
  • Agent frameworks: Deep hands-on experience with at least two of: Strands Agents SDK, LangGraph, CrewAI, AutoGen, LlamaIndex. Comfortable dropping to a custom orchestration loop where a framework gets in the way.
  • RAG engineering: Production experience with document processing and chunking strategy, embedding model selection, vector and hybrid (BM25 + dense) retrieval, reranking, query decomposition, GraphRAG and agentic RAG patterns, context-window budgeting, and grounding/citation enforcement. Able to diagnose whether a bad answer came from retrieval, ranking, or generation.
  • AWS AI stack – 3 years of experience:
  • Amazon Bedrock — Converse API, model selection and routing, Knowledge Bases, Guardrails, Flows, prompt caching, batch vs. real-time inference
  • Amazon SageMaker AI for custom model hosting, training, and endpoint operations
  • Vector and search: Amazon OpenSearch Serverless, S3 Vectors, Aurora PostgreSQL with pgvector, Amazon Kendra.
  • AWS core services – 5 years of experience: Lambda, Step Functions, ECS/EKS/Fargate, API Gateway, EventBridge, SQS, DynamoDB, S3, CloudWatch, IAM, KMS, Secrets Manager, VPC and PrivateLink. Able to design a VPC-isolated, private-endpoint deployment for a regulated client without help.
  • Python engineering – 7+ years of experience: Production-grade Python — async, typing, Pydantic, FastAPI, structured testing. Clean, reviewable code; not notebook-only.
  • Evaluation and LLMOps – 2+ years of experience: Offline and online evaluation design, RAGAS or equivalent retrieval metrics, trace-based observability (OpenTelemetry GenAI conventions, CloudWatch, Langfuse/LangSmith or similar), prompt and model versioning, token and cost governance, drift and regression detection.
  • AI security and responsible AI – 2+ years of experience: Prompt injection and tool-abuse threat modeling, scoped tool permissions and action approval design, data classification and PII handling, content filtering, auditability. Working familiarity with an AI governance framework such as NIST AI RMF.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Technical Architect - AWS Bedrock + Generative AI
Technical Architect - AWS Bedrock + Generative AI

WinWire Technologies Inc. • Hyderabad, Bengaluru

Hybrid
INR 3,000,000 - 6,000,000
Agentic AI Engineer
Agentic AI Engineer

EY • Hyderabad, Bengaluru, Delhi

On-site
INR 6,000,000 - 9,000,000
AWS Bedrock & Agentic AI Architect / Lead
AWS Bedrock & Agentic AI Architect / Lead

Zettamine Labs • Hyderabad

On-site
INR 1,800,000 - 3,000,000
Agentic AI with Python
Agentic AI with Python

Vibehackers • Coimbatore District

On-site
INR 4,000,000 - 9,000,000
Agentic AI Solutions Architect
Agentic AI Solutions Architect

Boundaryless • Pune District

On-site
INR 1,800,000 - 2,500,000
Principal AI Engineer
Principal AI Engineer

Willware Technologies • Bengaluru

On-site
INR 3,000,000 - 6,000,000
GenAI / Agentic AI Automation
GenAI / Agentic AI Automation

TymblHub • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior Agentic AI Engineer
Senior Agentic AI Engineer

Alternative Path • Gurgaon

On-site
INR 1,400,000 - 2,400,000
AI Engineer
AI Engineer

BT Group • Bengaluru

Hybrid
INR 2,500,000 - 4,500,000
Architect AI Data Engineer
Architect AI Data Engineer

EXL • Maharashtra

On-site
INR 3,000,000 - 4,200,000