Agentic AI Engineer - RAG Architecture - LLM Systems

AspenView Technology Partners

Bogotá

Híbrido

COP 381.667.000 - 572.501.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Competitive base
Benefits & wellness
Flexible work model
Growth opportunities
Inclusive culture

Descripción de la vacante

AspenView Technology Partners seeks a senior AI engineer to architect autonomous agentic workflows, design robust RAG pipelines, and evaluate LLMs for healthcare tasks. You will optimize vector databases, implement safety guardrails, and collaborate with US researchers and product teams to deliver compliant, scalable AI solutions.

Join a nearshore, remote-first team that values robust engineering, strong communication, and nearshore collaboration with US clients.

Formación

  • 4+ years in software production engineering, with at least 2+ years dedicated to building LLM applications, agentic workflows, and RAG architectures in production.

Responsabilidades

  • Architect and deploy autonomous agentic AI workflows and multi-agent systems.
  • Design, optimize, and scale production RAG pipelines with hybrid search, re-ranking, and context handling.
  • LLM fine-tuning and evaluation for healthcare tasks, establishing rigorous evaluation frameworks.
  • Manage vector stores for high-concurrency, low-latency similarity searches.
  • Implement enterprise safety controls and HIPAA compliance guardrails.
  • Collaborate with US clinical researchers, cloud architects, and product leads.

Conocimientos

Agentic frameworks
LangGraph
LlamaIndex
AutoGen
CrewAI
Python
LLM orchestration
Vector databases
Pinecone
Qdrant
Milvus
Weaviate
pgvector
TruLens

Herramientas

Pinecone
Qdrant
Milvus
Weaviate
pgvector
TruLens

Descripción del empleo

Build the Future with AspenView Technology Partners

At AspenView, we are passionate about transforming the way organizations approach technology. We specialize in creating high-performing, nearshore IT teams to help North American clients innovate faster and more efficiently.

As we continue to grow, we’re looking for exceptional people to join our team and help drive impactful change across industries.

Why Join AspenView?

At AspenView, we’re more than a nearshore IT partner— we’re a people-first, purpose-driven company that believes great culture drives great outcomes. We’re passionate about connecting talent and technology to deliver measurable value for clients—and meaningful career paths for our people.

Here’s what you can expect:
  • Competitive base
  • Comprehensive benefits and wellness support
  • Flexible work model: hybrid, remote, or in-office
  • Real growth opportunities and leadership visibility
  • Inclusive, respectful culture that blends U.S. innovation with Colombian heart
  • A company that listens, invests in you, and celebrates wins together
About The Role

In this role, you will lead the architecture and implementation of multi-step agentic workflows capable of reasoning, tool execution, and synthesizing complex clinical and operational data. You will engineer sophisticated RAG pipelines that bridge unstructured medical knowledge with enterprise EHR systems, establishing strict safety guardrails, hallucination controls, and HIPAA-compliant AI pipelines.

What You Will Do
  • Agentic System Engineering: Architect and deploy autonomous agentic AI workflows and multi-agent systems (using frameworks like LangGraph, LlamaIndex, AutoGen, or CrewAI) capable of multi-step reasoning, dynamic tool usage, and clinical task execution.
  • Advanced RAG Architecture: Design, optimize, and scale production RAG pipelines utilizing hybrid search (dense + sparse retrieval), re-ranking, query transformation, context compression, and semantic routing over multi-modal clinical and research datasets.
  • LLM Fine-Tuning & Evaluation: Evaluate, fine-tune, and benchmark foundational models (e.g., Llama 3, Claude, GPT-4, Med-PaLM) for specialized healthcare tasks, establishing rigorous evaluation frameworks (e.g., Ragas, TruLens) for accuracy, groundness, and latency.
  • Vector Database Infrastructure: Manage and optimize vector stores (Pinecone, Qdrant, Milvus, Weaviate, or pgvector) for high-concurrency, low-latency similarity searches and knowledge retrieval.
  • Guardrails & HIPAA Compliance: Implement enterprise safety controls, input/output sanitization, and hallucination guardrails (e.g., NeMo Guardrails, Llama Guard) to ensure 100% HIPAA compliance and zero unauthorized PHI leakage.
  • Cross-Functional Innovation: Collaborate closely with US-based clinical researchers, cloud architects, and product leads to translate complex healthcare needs into autonomous AI capabilities.
Experience
What You Bring
  • 4+ years in software production engineering, with at least 2+ years dedicated to building LLM applications, agentic workflows, and RAG architectures in production.
Technical Expertise
  • Agentic Frameworks & Orchestration: Strong mastery of modern AI orchestration ecosystems (LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI).
  • Search & Retrieval Infrastructure: Deep knowledge of Vector Databases (Pinecone, Qdrant, Milvus, pgvector), embedding models, semantic chunking strategies, and re-ranking models (e.g., Cohere Rerank, BGE).
  • Programming & Cloud AI Stack: Advanced proficiency in Python, with hands‑on deployment experience on cloud platforms (AWS Bedrock / SageMaker or Azure OpenAI / Vertex AI).
  • Model Evaluation & Guardrails: Experience implementing LLM observability, evaluation metrics (faithfulness, answer relevance), and safety frameworks.
  • Healthcare Interoperability (Preferred): Familiarity with medical ontologies (SNOMED, LOINC, ICD-10) and healthcare data interfaces (FHIR APIs, HL7) is a strong plus.
Language Proficiency
  • Advanced/Fluent English (C1/C2) for daily technical collaboration with Boston-based technology and research leadership.
Soft Skills & Competencies
  • Systems Thinking for Non-Deterministic AI: Exceptional ability to debug, evaluate, and stabilize non-deterministic model outputs for mission-critical healthcare applications.
  • Clinical & User Empathy: Deep commitment to designing transparent, explainable, and safe AI systems that empower clinicians rather than replace them.
  • Proactive Nearshore Ownership: Self-starter capable of driving end-to-end AI features autonomously in a remote-first, agile nearshore setting.
  • Clear Technical Storytelling: Ability to communicate complex AI concepts, architectural trade-offs, and evaluation metrics clearly to non-technical stakeholders.
Visa Sponsorship

AspenView does not sponsor employment visas for this role. Applicants must be currently authorized to work in the United States on a permanent basis without the need for visa sponsorship now or in the future.

Equal Opportunity Employer

AspenView is proud to be an equal opportunity employer. We believe in creating an environment where all employees feel welcome, valued, and empowered to succeed. We celebrate diversity and strive to build a culture of inclusion where all individuals, regardless of their race, color, gender, gender identity or expression, sexual orientation, disability, age, or any other characteristic, can thrive. We encourage applicants from all walks of life to join our team and make a lasting impact.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Principal Enterprise Architect – AI & Cloud
Principal Enterprise Architect – AI & Cloud

AspenView Technology Partners • Bogotá

Híbrido
COP 381.667.000 - 572.501.000
Competitive base
Comprehensive benefits and wellness
Flexible work model: hybrid/remote/in‑
+2
Remote Agentic AI Engineer: RAG Architect for LLM Systems
Remote Agentic AI Engineer: RAG Architect for LLM Systems

AspenView Technology Partners • Bogotá

Híbrido
COP 381.667.000 - 572.501.000
Competitive base
Benefits & wellness
Flexible work model
+2
AWS Solutions Architect – AI, Blockchain & Computational Mathematics
AWS Solutions Architect – AI, Blockchain & Computational Mathematics

AspenView Technology Partners • Bogotá

Híbrido
COP 90.000.000 - 130.000.000
Competitive base
Benefits & wellness
Flexible work model
+2
Senior Data Engineer – AWS Infrastructure
Senior Data Engineer – AWS Infrastructure

AspenView Technology Partners • Bogotá

Híbrido
COP 286.250.000 - 381.667.000
Competitive base
Benefits and wellness
Flexible work model
+2
Senior AI Engineer
Senior AI Engineer

Athenaworks • Bogotá

Presencial
COP 388.878.000 - 583.318.000
Payment in USD
Flexible work schedule
Non-working pay days
+1
GTM Engineer
GTM Engineer

AspenView Technology Partners, Inc. • Colombia

Híbrido
COP 288.258.000 - 448.402.000
Competitive base
Benefits & wellness
Flexible work model
+2
Lead AI/ML Engineer
Lead AI/ML Engineer

EPAM Systems • Colombia

Presencial
COP 282.380.000 - 376.506.000
Healthcare benefits
Global career opportunities
Paid time off
AI Engineer (Generative AI / Machine Learning)
AI Engineer (Generative AI / Machine Learning)

Golabs Tech • Colombia

A distancia
COP 187.447.000 - 374.895.000
Competitive Compensation
100% Remote Work
Paid Time Off
+5
Senior AI Engineer: LLMs, RAG & Cloud-Native Systems
Senior AI Engineer: LLMs, RAG & Cloud-Native Systems

Xebia • Bogotá

Presencial
AI Software Engineer
AI Software Engineer

Vichara Technologies, Inc. • Fagua

Presencial
COP 1.200.000 - 1.500.000