Senior Cloud DevOps Engineer (AWS, Kubernetes, Argo)

Agileengine

Xico

Hybrid

MXN 2,172,000 - 3,258,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Professional growth
Competitive compensation: USD-based
Exciting projects
Flextime

Job summary

AgileEngine is seeking a Senior Cloud/DevOps Engineer to operate and improve the cloud infrastructure and reliability layers behind an enterprise data platform in a regulated healthcare environment. The role requires 5+ years in Cloud Engineering or SRE, advanced hands-on AWS/EKS and Argo Workflows skills, and strong English communication.

The candidate will lead incident resolution, implement IaC standards with Terraform, and enhance observability across data workloads.

Qualifications

  • 5+ years of professional experience in Cloud Engineering, DevOps or Site Reliability Engineering.
  • Strong hands-on experience operating AWS infrastructure in production environments.
  • Advanced experience with Kubernetes and Amazon EKS, including workload operations, troubleshooting, access, observability, capacity, and reliability.
  • Hands-on experience administering and troubleshooting Argo Workflows or comparable workflow orchestration platforms.
  • Strong Infrastructure as Code experience with Terraform and source-controlled infrastructure practices.
  • Experience building, hardening, and supporting CI/CD pipelines and production release processes.
  • Strong experience with monitoring, logging, alerting, and incident-routing tools such as Splunk, PagerDuty, Opsgenie, or comparable platforms.
  • Demonstrated ability to lead complex incident resolution, perform root-cause analysis, and translate findings into preventive improvements.
  • Proficiency in automation and scripting using Python, Shell, Bash, or similar languages.
  • Ability to make well-reasoned technical decisions, identify tradeoffs, estimate work, and drive improvements across a complex platform.
  • Experience mentoring engineers and collaborating effectively with Data Engineering, Security, Governance, Analytics, and business stakeholders.
  • Strong written and verbal English communication skills, with the ability to work directly with client stakeholders.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed on-call rotation.

Responsibilities

  • Provide senior technical ownership for the Cloud / DevOps service tower during the LatAm coverage window, including day-to-day operations, complex troubleshooting, and L2/L3 escalation.
  • Operate, maintain, and improve AWS infrastructure supporting the Data Platform, including Amazon EKS, S3, EventBridge, SQS, API Gateway, Lambda, and related services.
  • Administer Kubernetes-hosted workloads and Argo Workflows, including deployment, scheduling, monitoring, troubleshooting, capacity management, resiliency, and recovery.
  • Define and improve standards for Infrastructure as Code, configuration management, CI/CD, release execution, rollback, and environment consistency, primarily using Terraform and Git-based delivery practices.
  • Lead the consolidation and improvement of observability across infrastructure and data workloads, linking alerts to operational evidence from Argo, dbt, Snowflake, and supporting runbooks.
  • Improve alert routing and escalation workflows across tools such as Splunk, Opsgenie, PagerDuty, Microsoft Teams, and data-specific observability platforms.
  • Lead or support major incident response, root-cause analysis, post-incident reviews, and corrective actions, with clear communication to technical and service stakeholders.
  • Design and implement reliability improvements such as selective auto-remediation, dependency-aware alert correlation, impact analysis, and automation of repetitive operational work.
  • Track and contribute to service metrics including availability, SLA compliance, alert volumes, workflow reliability, deployment outcomes, and mean time to restore service.
  • Apply disciplined change-management, access-control, secrets-management, auditability, and documentation practices appropriate for a HIPAA-, GDPR-, and FDA-regulated environment.
  • Create and maintain runbooks, operating procedures, architecture context, recovery procedures, and knowledge-transfer materials.
  • Mentor Middle-level engineers, review technical work, improve team practices, and promote consistent execution across the distributed team.
  • Participate in the Cloud / DevOps on-call rotation for critical incidents outside staffed service hours.

Skills

5+ years
AWS
Kubernetes
Amazon EKS
Argo Workflows
Terraform
CI/CD pipelines
Monitoring
Incident management
Python scripting
On-call

Tools

Argo Workflows
Terraform
Kubernetes
Jira

Job description

AgileEngine is seeking a Senior Cloud/DevOps Engineer to operate and improve the cloud infrastructure and reliability layers behind an enterprise data platform in a regulated healthcare environment. The role requires 5+ years in Cloud Engineering or SRE, advanced hands-on AWS/EKS and Argo Workflows skills, and strong English communication.

The candidate will lead incident resolution, implement IaC standards with Terraform, and enhance observability across data workloads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud/DevOps Engineer — AWS, Kubernetes & Argo
Senior Cloud/DevOps Engineer — AWS, Kubernetes & Argo

AgileEngine, LLC. • Santiago de Querétaro

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive pay (USD)
Exciting projects
+1
Senior Cloud/DevOps Engineer - Remote AWS, Kubernetes & Argo
Senior Cloud/DevOps Engineer - Remote AWS, Kubernetes & Argo

AgileEngine, LLC. • Región Centro

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive pay
Exciting projects
+1
Senior Cloud & DevOps Engineer — AWS, Kubernetes, Argo
Senior Cloud & DevOps Engineer — AWS, Kubernetes, Argo

AgileEngine, LLC. • Monterrey

On-site
MXN 1,200,000 - 2,000,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Senior Cloud & DevOps Engineer — Remote & Flexible
Senior Cloud & DevOps Engineer — Remote & Flexible

AgileEngine, LLC. • León

On-site
MXN 900,000 - 1,300,000
Professional growth
USD-based pay
Exciting projects
+1
Senior Cloud & DevOps Engineer - Remote & Flexible
Senior Cloud & DevOps Engineer - Remote & Flexible

AgileEngine, LLC. • Rosarito

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer: AWS, Kubernetes, Terraform
Senior Cloud/DevOps Engineer: AWS, Kubernetes, Terraform

AgileEngine, LLC. • Puebla de Zaragoza

On-site
MXN 900,000 - 1,300,000
Remote Cloud Platform Lead — AWS, Kubernetes & Argo
Remote Cloud Platform Lead — AWS, Kubernetes & Argo

AgileEngine, LLC. • Puebla de Zaragoza

On-site
MXN 2,172,000 - 3,258,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Remote Cloud Platform Lead (AWS/K8s)
Remote Cloud Platform Lead (AWS/K8s)

AgileEngine, LLC. • Rosarito

On-site
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud Platform Lead - AWS, Kubernetes, Data Platform
Senior Cloud Platform Lead - AWS, Kubernetes, Data Platform

AgileEngine, LLC. • León

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Rosarito

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1