Senior AI/ML Full-Stack Engineer: RAG & LLM Orchestration

MAS Global Consulting

Plano (TX)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MAS Global Consulting is seeking a Senior AI/ML Full Stack Engineer to design and deploy production-grade AI applications. You will build end-to-end systems, including RAG pipelines, vector stores, and AI agents, using Java or Python in a fast-paced, in-office environment in the US.

You will work with LangChain, LlamaIndex, Semantic Kernel, and CrewAI, deploying on AWS Bedrock and Claude models, with strong emphasis on guardrails and responsible AI practices.

Qualifications

  • Bachelor's degree required.
  • 10+ years of software development experience (Java or Python), OR 7+ years if entirely full-stack + Agentic AI development experience.
  • 2+ years hands-on experience building AI/ML applications in production.
  • Strong proficiency with RAG architectures — chunking strategies, embedding models, vector stores.
  • Experience with AI orchestration frameworks: LangChain, LlamaIndex, Semantic Kernel, or CrewAI.
  • Hands-on experience with AWS Bedrock, Anthropic Claude models, and model invocation APIs.
  • Proven prompt engineering skills — system prompts, few-shot, chain-of-thought, tool use, structured outputs.
  • Experience building conversational AI: chatbots (text) and voicebots.
  • Proficiency with AWS services (Lambda, Step Functions, API Gateway, S3, DynamoDB, SQS).
  • Experience with CI/CD pipelines, containerization (Docker, ECS/EKS), and infrastructure-as-code.
  • Strong understanding of API design (REST, GraphQL), microservices architecture, and event-driven systems.
  • Familiarity with evaluation frameworks for LLM outputs.
  • Experience with guardrails, content filtering, and responsible AI practices

Responsibilities

  • Design and build AI/ML applications end-to-end, from architecture through production deployment.
  • Implement RAG pipelines, including chunking strategies, embedding models, and vector store integration.
  • Build and orchestrate AI agents using frameworks such as LangChain, LlamaIndex, Semantic Kernel, or CrewAI.
  • Develop and deploy solutions using AWS Bedrock, Anthropic Claude models, and model invocation APIs.
  • Apply advanced prompt engineering techniques — system prompts, few-shot, chain-of-thought, tool use, structured outputs.
  • Build conversational AI experiences, including chatbots (text) and voicebots (speech-to-text, text-to-speech).
  • Design and maintain APIs (REST, GraphQL), microservices, and event-driven architectures.
  • Own CI/CD pipelines, containerization (Docker, ECS/EKS), and infrastructure-as-code for production systems.
  • Implement evaluation frameworks, guardrails, content filtering, and responsible AI practices across LLM-powered features.

Skills

AI/ML full stack development
Prompt engineering
Agentic AI development
Production-grade software development
Hands-on problem solving

Education

Bachelor's degree

Tools

Pinecone
OpenSearch
pgvector
FAISS
LangChain
LlamaIndex
Semantic Kernel
CrewAI
AWS Bedrock
Anthropic Claude
Docker
ECS/EKS
API Gateway
Lambda
S3
DynamoDB
SQS

Job description

MAS Global Consulting is seeking a Senior AI/ML Full Stack Engineer to design and deploy production-grade AI applications. You will build end-to-end systems, including RAG pipelines, vector stores, and AI agents, using Java or Python in a fast-paced, in-office environment in the US.

You will work with LangChain, LlamaIndex, Semantic Kernel, and CrewAI, deploying on AWS Bedrock and Claude models, with strong emphasis on guardrails and responsible AI practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Full-Stack Engineer: Scale with LLMs & RAG
Senior AI Full-Stack Engineer: Scale with LLMs & RAG

ASAL TECHNOLOGIES • Palestine (TX)

On-site
USD 120,000 - 180,000
Full Stack AI Engineer
Full Stack AI Engineer

MAS Global Consulting • Plano (TX)

On-site
USD 140,000 - 190,000
Enterprise AI Engineer: RAG & LLM Solutions
Enterprise AI Engineer: RAG & LLM Solutions

Stack AI, Inc. • San Francisco (CA)

On-site
USD 110,000 - 150,000
Senior AI Architect for LLMs, RAG & Enterprise AI
Senior AI Architect for LLMs, RAG & Enterprise AI

Ll Oefentherapie • Oklahoma City (OK)

On-site
USD 140,000 - 230,000
Senior AI Engineer: RAG & LLM Stack
Senior AI Engineer: RAG & LLM Stack

Fanisko • United States

On-site
USD 120,000 - 160,000
Senior AI Systems Engineer – RAG, LLMs & Multi-Agent Orchestration
Senior AI Systems Engineer – RAG, LLMs & Multi-Agent Orchestration

Tallgrass • Lakewood (CO)

On-site
USD 106,000 - 131,000
Health insurance
401(k) with match
Annual bonus
+2
RAG ML Engineer — Retrieval & LLM Orchestration
RAG ML Engineer — Retrieval & LLM Orchestration

S&P Global, Inc. • New York (NY)

On-site
USD 140,000 - 180,000
Medical, Dental, and Vision insurance
Unlimited Paid Time Off
Parental Leave (26 weeks)
Senior AI/ML Engineer: LLM, RAG & End-to-End Deployment
Senior AI/ML Engineer: LLM, RAG & End-to-End Deployment

Vizient, Inc. • Centennial (CO)

On-site
USD 102,000 - 179,000
Incentive eligible
Benefits plan
AI Engineer — RAG & LLM Systems, Multi-Agent
AI Engineer — RAG & LLM Systems, Multi-Agent

VDart • Frisco (TX)

On-site
USD 131,000 - 200,000
AI Data Engineer: LLMs, RAG & Multi-Agent Systems
AI Data Engineer: LLMs, RAG & Multi-Agent Systems

Crossing Hurdles • United States

On-site
USD 140,000 - 190,000