AI Software Engineer

Vichara Technologies, Inc.

Fagua

Presencial

COP 1.200.000 - 1.500.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Vichara Technologies, Inc. is looking for an experienced AI/ML Engineer to architect and lead multi-agent LLM systems. This role involves building Retrieval-Augmented Generation pipelines and integrating various models to enhance task-specific Q&A functionalities.

Ideal candidates will have 5+ years of experience in AI/ML, with required skills in multi-agent orchestration and RAG frameworks. The position is based in Colombia and offers a competitive salary package.

Formación

  • 5+ years as an AI or ML Engineer.
  • Experience with RAG benchmark automation.
  • Familiarity with Credit Risk, Banking, and Investment Analytics.

Responsabilidades

  • Architect, design, and lead multi-agent LLM systems.
  • Build Retrieval-Augmented Generation (RAG) pipelines.
  • Define system workflows for summarization and query routing.
  • Integrate GPT-4o, PaLM 2, and other models for contextual Q&A.
  • Manage prompt routing and variant testing.

Conocimientos

RAG Frameworks: LanceDB, Pinecone, ElasticSearch, FAISS, MongoDB
Agentic AI: LangGraph multi-agent orchestration, routing logic
Fine-Tuning: BERT / domain-specific transformer tuning
Reranker-based retrieval knowledge (MiniLM / CrossEncoder)
Prompt evaluation and scoring knowledge (BLEU, ROUGE)

Herramientas

MongoDB
Redis Streams
OpenTelemetry
AWS EKS / Azure Kubernetes Service
CI/CD pipelines (Azure DevOps)

Descripción del empleo

  • Compensation: COP 1200000 - COP 1500000 - yearly
Company Description

Vichara is a Financial Services focused products and services firm headquartered in NY and building systems for some of the largest i-banks and hedge funds in the world.

Job Description

Key Responsibilities

Architect, design, and lead multi-agent LLM systems using LangGraph, LangChain, and Promptfoo for prompt lifecycle management and benchmarking.

Build Retrieval-Augmented Generation (RAG) pipelines leveraging hybrid vector search (dense + keyword) using LanceDB, Pinecone, or Elasticsearch.

Define system workflows for summarization, query routing, retrieval, and response generation, ensuring minimal latency and high precision.

Develop RAG evaluation frameworks combining retrieval precision/recall, hallucination detection, and latency metrics — aligned with analyst and business use cases.

Integrate GPT-4o, PaLM 2, and open-weight models (LLaMA, Mistral) for task-specific contextual Q&A.

Fine-tune transformer models (BERT, SentenceTransformers) for document classification, summarization, and sentiment analysis.

Manage prompt routing and variant testing using Promptfoo or equivalent tools.

Implement multi-agent architectures with modular flows — enabling task-specific agents for summarization, retrieval, classification, and reasoning.

Design fallback and recovery behaviors to ensure robustness in production.

Employ LangGraph for parallel and stateful agent orchestration, error recovery, and deterministic flow control.

Architect ingestion pipelines for structured and unstructured data — including financial statements, filings, and PDF documents.

Leverage MongoDB for metadata storage and Redis Streams for async task execution and caching.

Implement vector-based search and retrieval layers for high-throughput and low-latency AI systems.

Observability & Production Deployment

Deploy end-to-end AI systems on AWS EKS / Azure Kubernetes Service, integrated with CI/CD pipelines (Azure DevOps).

Build comprehensive monitoring dashboards using OpenTelemetry and Signoz, tracking latency, retrieval precision, and application health.

Enforce testing and regression validation using golden datasets and structured assertion checks for all LLM responses.

Collaborate with DevOps, MLOps, and application development teams to integrate AI APIs with React / FastAPI-based user interfaces.

Work with business analysts to translate credit, compliance, and customer-support requirements into actionable AI agent workflows.

Mentor a small team of GenAI developers and data engineers in RAG, embeddings, and orchestration techniques.

Qualifications
  • Experience:
    • 5+ years as an AI or ML Engineer

Required Skills & Experience

RAG Frameworks: LanceDB, Pinecone, ElasticSearch, FAISS, MongoDB

Agentic AI: LangGraph multi-agent orchestration, routing logic, task decomposition

Fine-Tuning: BERT / domain-specific transformer tuning, evaluation framework design

Knowledge of Reranker-based retrieval (MiniLM / CrossEncoder)

Familiarity with Prompt evaluation and scoring (BLEU, ROUGE, Faithfulness)

Domain exposure to Credit Risk, Banking, and Investment Analytics

Experience with RAG benchmark automation and model evaluation dashboards

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior AI Engineer: LLMs, RAG & Cloud-Native Systems
Senior AI Engineer: LLMs, RAG & Cloud-Native Systems

Xebia • Bogotá

Presencial
Senior AI Engineer
Senior AI Engineer

Intellias • Colombia

Presencial
COP 120.000.000 - 180.000.000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Xebia • Bogotá

Presencial
INT Agentic AI Engineer
INT Agentic AI Engineer

Aditi Consulting • Bogotá

Presencial
COP 120.000.000 - 160.000.000
AI Engineer (Generative AI / Machine Learning)
AI Engineer (Generative AI / Machine Learning)

Golabs Tech • Colombia

A distancia
COP 187.447.000 - 374.895.000
Competitive Compensation
100% Remote Work
Paid Time Off
+5
AI Engineer (Generative AI | RAG | AI Agents)
AI Engineer (Generative AI | RAG | AI Agents)

10x Advisory • Colombia

Presencial
COP 90.000.000 - 180.000.000
Senior IA Engineer (Python + AWS)
Senior IA Engineer (Python + AWS)

GlobalLogic • Colombia

Presencial
COP 120.000.000 - 180.000.000
Senior AI Engineer
Senior AI Engineer

Athenaworks • Bogotá

Presencial
COP 388.878.000 - 583.318.000
Payment in USD
Flexible work schedule
Non-working pay days
+1
AI Engineer
AI Engineer

Auxis • Bogotá

Presencial
COP 288.258.279 - 416.373.070
Professional development opportunities
Competitive compensation
Senior AI Engineer
Senior AI Engineer

Crew • Colombia

Presencial
COP 182.302.000 - 255.223.000