Technical Lead

Tata Consultancy Services

Cary (NC)

On-site

USD 110,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tata Consultancy Services in Cary, NC is seeking a senior software engineer to architect and productionize cloud native backend services and AI/LLM inference pipelines.

You will design Python-based APIs and microservices with FastAPI, implement LLM capabilities, and ensure observability, CI/CD, and scalable deployments using Kubernetes and Helm. This role emphasizes performance, reliability, and collaboration with frontend teams.

Qualifications

  • Experience building cloud native backend services and AI/LLM inference pipelines.
  • Design and develop Python-based APIs and microservices (FastAPI, async patterns).
  • Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and versioning.
  • Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.

Responsibilities

  • Build and productionize cloud native backend services and AI/LLM inference pipelines.
  • Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.
  • Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.
  • Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.
  • Build event driven, resilient integrations and containerized services, with hands‑on Kubernetes debugging and Helm-based deployments.
  • Establish observability, SLOs, CI/CD automation, testing.

Skills

Cloud native backends
Python APIs
LangChain/LangGraph
LLM workflows
Kubernetes
Helm deployments
CI/CD observability

Education

Bachelor of Computer Science

Tools

FastAPI
Async patterns
LangChain
LangGraph
Azure/AKS

Job description

Job Description
  • Build and productionize cloud native backend services and AI/LLM inference pipelines.
  • Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.
  • Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.
  • Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.
  • Build event driven, resilient integrations and containerized services, with hands‑on Kubernetes debugging and Helm-based deployments.
  • Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.
  • Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.
Must Have Technical/Functional Skills
  • 13+ years of experience with IT
  • Build and productionize cloud native backend services and AI/LLM inference pipelines.
  • Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.
  • Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.
  • Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.
  • Build event driven, resilient integrations and containerized services, with hands‑on Kubernetes debugging and Helm-based deployments.
  • Establish observability, SLOs, CI/CD automation, testing
  • Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.
  • Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.
Roles & Responsibilities
  • Build and productionize cloud native backend services and AI/LLM inference pipelines.
  • Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.
  • Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.
  • Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.
  • Build event driven, resilient integrations and containerized services, with hands‑on Kubernetes debugging and Helm-based deployments.
  • Establish observability, SLOs, CI/CD automation, testing
  • Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.
  • Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.
Salary Range

Salary Range: $110,000 to $130,000 per year

Qualifications

BACHELOR OF COMPUTER SCIENCE

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Solution Lead/ JAVA TECH LEAD
Solution Lead/ JAVA TECH LEAD

Connexions • Carolina Beach (NC)

On-site
USD 150,000 - 210,000
Python AI Solutions
Python AI Solutions

Centraprise • Cary (NC)

On-site
USD 140,000 - 190,000
AI/ML Engineer
AI/ML Engineer

Jobtailor • Colorado

On-site
USD 120,000 - 160,000
AIML Engineer
AIML Engineer

Tata Consultancy Services • Milford (OH)

On-site
USD 140,000 - 165,000
Discretionary Annual Incentive
Comprehensive Medical Coverage: Health
Dental & Vision
+4
Sr AI/ML Engineer
Sr AI/ML Engineer

Vizient • Irving (TX)

On-site
USD 102,000 - 179,000
AI Engineer
AI Engineer

BayRockLabs • Newark (CA)

Hybrid
Tech Lead Manager
Tech Lead Manager

ChenMed • Town of Florida (NY)

On-site
USD 118,000 - 169,000
Lead Backend Architect
Lead Backend Architect

Veriipro • New York (NY)

On-site
USD 180,000 - 260,000
Lead Backend Architect: AI-Driven Cloud-Native Systems
Lead Backend Architect: AI-Driven Cloud-Native Systems

Veriipro • New York (NY)

On-site
USD 180,000 - 260,000
Senior AI EngineerWoodland Hills
Senior AI EngineerWoodland Hills

TechDigital Group • Los Angeles (CA)

On-site
USD 120,000 - 160,000