MLOps Engineer

XenonStack Moments

Vallecito 15-8 Hualtaco I

Presencial

PEN 90.000 - 130.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura hecha para este puesto de trabajo — un currículum y una carta de presentación adaptados que responden directamente a la oferta.

Supera los filtros ATS

Descripción de la vacante

XenonStack is seeking an MLOps Engineer to deploy, monitor, and optimize agentic AI systems in production environments across cloud and on-prem. You will own observability, reliability, and integration with enterprise APIs, knowledge bases, and third-party systems, working at the intersection of MLOps, DevOps, and AI.

Join a fast-growing AI Foundry building scalable, secure platforms for enterprise agents, and help ensure reliable, explainable AI with strong governance and performance monitoring.

Formación

  • 2–5 years of experience in DevOps, MLOps, or AI systems engineering.
  • Hands-on with LLM orchestration frameworks.
  • Familiarity with AgentOps tools like LangSmith, PromptLayer, Weights & Biases, Arize AI.
  • Proficiency in Python and scripting for automation.
  • Experience with cloud platforms (AWS, GCP, Azure) and containerization (Docker, Kubernetes).
  • Knowledge of monitoring & observability tools (Prometheus, Grafana, ELK, OpenTelemetry).
  • Understanding of RAG pipelines, vector databases, and context orchestration.

Responsabilidades

  • Agent Deployment & Operations across cloud and on-prem environments.
  • Integrate agents with enterprise APIs, knowledge bases, and third-party systems.
  • Implement agentic observability frameworks to track performance, latency, cost, and accuracy.
  • Monitor execution traces, context windows, and agent interactions for anomalies.
  • Fine-tune agent configurations for cost efficiency, scalability, and response quality.
  • Implement fallbacks, guardrails, and redundancy mechanisms for reliability.
  • Build automated pipelines to evaluate agent performance, safety, and compliance.
  • Feed test results into continuous improvement loops with ML/AI engineers.
  • Ensure agents follow enterprise security, governance and compliance standards.
  • Collaborate with Responsible AI teams on trust, safety, and audit mechanisms.
  • Work with AI engineers, DevOps, and product managers to align reliability with business outcomes.
  • Support customer deployments with customized monitoring and tuning.

Conocimientos

LLM orchestration
Python
Cloud platforms
Docker
Kubernetes
LangChain
Observability tools
AgentOps tools
Monitoring

Herramientas

LangChain
LangGraph
LlamaIndex
Prometheus
Grafana
ELK
OpenTelemetry
Arize AI

Descripción del empleo

About Xenonstack

XenonStack is a Data and AI Foundry for Agentic Systems, enabling enterprises to design, deploy, operate, and scale intelligent agents across digital and physical environments.

We Build Enterprise-grade Platforms Across The Agentic Stack
  • Akira AI - Reasoning and agent orchestration. Turn models into collaborative, policy-governed agents that learn and act together.
  • ElixirData - Agentic analytics intelligence. Explainable, decision-centric analytics for measurable business outcomes.
  • NexaStack - Agentic infrastructure automation. Secure, compliant AI deployment across cloud, edge, and on-prem.
  • MetaSecure - Trust, compliance and defense. Continuous assurance with AI-BOMs, risk scoring, and agentic security.

Our mission is to accelerate the world’s transition to AI + Human Intelligence by making agentic systems reliable, responsible, and enterprise-ready.

The Opportunity

We are seeking an MLOps Engineer to deploy, monitor, and optimize agentic AI systems in production environments.

This role is at the core of AI observability and operational reliability, ensuring that multi-agent workflows perform consistently, safely, and efficiently. You will work at the intersection of MLOps, DevOps, and Agentic AI, enabling enterprises to confidently adopt AI agents at scale.

Key Responsibilities
  • Agent Deployment & Operations
    • Deploy and maintain LLM-powered and multi-agent systems across cloud and on-prem environments.
    • Integrate agents with enterprise APIs, knowledge bases, and third-party systems.
  • Monitoring & Observability
    • Implement agentic observability frameworks to track performance, latency, cost, and accuracy.
    • Monitor execution traces, context windows, and agent interactions for anomalies.
  • Optimization & Reliability
    • Fine-tune agent configurations for cost efficiency, scalability, and response quality.
    • Implement fallbacks, guardrails, and redundancy mechanisms to ensure reliability.
  • Evaluation & Feedback Loops
    • Build automated pipelines to evaluate agent performance, safety, and compliance.
    • Feed test results into continuous improvement loops with ML/AI engineers.
  • Security & Compliance
    • Ensure agents follow enterprise security, governance, and compliance standards.
    • Collaborate with Responsible AI teams to implement trust, safety, and audit mechanisms.
  • Cross-Functional Collaboration
    • Work with AI engineers, DevOps, and product managers to align operational reliability with business outcomes.
    • Support customer deployments with customized monitoring and tuning.
Skills & Qualifications
Must-Have
  • 2–5 years of experience in DevOps, MLOps, or AI systems engineering.
  • Hands-on with LLM orchestration frameworks (LangChain, LangGraph, LlamaIndex).
  • Familiarity with AgentOps tools (LangSmith, PromptLayer, Weights & Biases, Arize AI).
  • Proficiency in Python and scripting for automation.
  • Experience with cloud platforms (AWS, GCP, Azure) and containerization (Docker, Kubernetes).
  • Knowledge of monitoring & observability tools (Prometheus, Grafana, ELK, OpenTelemetry).
  • Understanding of RAG pipelines, vector databases, and context orchestration.
Good-to-Have
  • Exposure to multi-agent orchestration (MCP, A2A messaging, AgentBridge).
  • Experience in Responsible AI / model evaluation frameworks.
  • Familiarity with CI/CD pipelines for AI models and agents.
  • Background in BFSI, GRC, SOC, or enterprise SaaS systems.
WHY SHOULD YOU JOIN US?
  • Agentic AI Product Company
  • A Fast-Growing Category Leader
  • Career Mobility & Growth
  • Global Exposure
  • Create Real Impact
  • Culture of Excellence
  • Responsible AI First

Work on next-gen AI platforms where agent reliability and observability define enterprise adoption.

Join one of the fastest-growing AI Foundries, powering mission-critical AI agents for global enterprises.

Grow into roles such as AgentOps Lead, Reliability Engineer, or AI Systems Architect.

Manage enterprise-scale AgentOps deployments across regulated industries worldwide.

Ensure that AI agents in production deliver measurable business outcomes.

Our values — Agency, Taste, Ownership, Mastery, Impatience, and Customer Obsession — empower you to build, innovate, and own outcomes.

Contribute to trustworthy, explainable, and compliant AI agents that enterprises can rely on.

XENONSTACK CULTURE – JOIN US & MAKE AN IMPACT!

At XenonStack, we believe in shaping the future of intelligent systems. We foster a culture of cultivation built on bold, human-centric leadership principles, where deep work, simplicity, and adoption define everything we do.

Our Cultural Values
  • Agency – Be self-directed and proactive.
  • Taste – Sweat the details and build with precision.
  • Ownership – Take responsibility for outcomes.
  • Mastery – Commit to continuous learning and growth.
  • Impatience – Move fast and embrace progress.
  • Customer Obsession – Always put the customer first.
Our Product Philosophy
  • Obsessed with Adoption – Making AI agents reliable and enterprise-ready.
  • Obsessed with Simplicity – Turning complex agent operations into seamless, intuitive workflows.

Be a part of our mission to accelerate the world’s transition to AI + Human Intelligence — by ensuring AI agents are reliable, secure, and production-ready.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

MLOps Engineer
MLOps Engineer

XenonStack Moments • San Martin Cp3 Hualtaco I

Presencial
PEN 60.000 - 100.000
Site Reliability Engineer, AI Platform
Site Reliability Engineer, AI Platform

XenonStack Moments • Vallecito 15-8 Hualtaco I

Presencial
PEN 308.000 - 444.000
Site Reliability Engineer, AI Platform
Site Reliability Engineer, AI Platform

XenonStack Moments • San Martin Cp3 Hualtaco I

Presencial
PEN 376.000 - 478.000
Machine Learning Engineer, Agentic Systems
Machine Learning Engineer, Agentic Systems

XenonStack Moments • San Martin Cp3 Hualtaco I

Presencial
PEN 204.000 - 305.000
Medical Insurance
Certifications
Recognition program
Machine Learning Engineer, Agentic Systems
Machine Learning Engineer, Agentic Systems

XenonStack Moments • Vallecito 15-8 Hualtaco I

Presencial
PEN 60.000 - 100.000
Continuous Learning
Certifications & Workshops
Cutting-edge Projects
+4
Forward Deployed Solutions Engineer, Agentic Systems
Forward Deployed Solutions Engineer, Agentic Systems

XenonStack Moments • San Martin Cp3 Hualtaco I

Híbrido
PEN 305.000 - 475.000
Hybrid work model
Global exposure
Career growth opportunities
Forward Deployed Solutions Engineer, Agentic Systems
Forward Deployed Solutions Engineer, Agentic Systems

XenonStack Moments • Vallecito 15-8 Hualtaco I

Híbrido
PEN 308.000 - 444.000
Machine Learning Engineer, Reinforcement Learning
Machine Learning Engineer, Reinforcement Learning

XenonStack Moments • Vallecito 15-8 Hualtaco I

Presencial
PEN 100.000 - 180.000
Machine Learning Engineer, Reinforcement Learning
Machine Learning Engineer, Reinforcement Learning

XenonStack Moments • San Martin Cp3 Hualtaco I

Presencial
PEN 305.000 - 407.000
Solution Architect, DevOps
Solution Architect, DevOps

XenonStack Moments • San Martin Cp3 Hualtaco I

Presencial
PEN 373.000 - 475.000
Comprehensive medical insurance
Leadership opportunities
Professional development budget