Remote Cloud Platform Lead — AWS, Kubernetes & Argo

AgileEngine, LLC.

Puebla de Zaragoza

On-site

MXN 2,172,000 - 3,258,000

Full time

8 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Professional growth
Competitive USD-based pay
Exciting projects
Flextime: remote/office options

Job summary

AgileEngine, LLC. is seeking a Cloud Platform Technical Lead in Puebla, Mexico to own the reliability and orchestration of an enterprise data platform in a regulated healthcare environment.

You will lead transition to managed services, define IaC and CI/CD standards, and oversee cross-functional incident response. The role requires deep AWS, Kubernetes/EKS, Argo Workflows, Terraform, Snowflake, S3 data lakes, and data pipeline experience, with strong English communication and LatAm hours

Qualifications

  • 6+ years of professional experience in Cloud Engineering, Platform Engineering, DevOps or Site Reliability Engineering.
  • Deep hands-on expertise with AWS production environments, including cloud architecture, security, networking fundamentals, access management, monitoring, capacity, and operational troubleshooting.
  • Advanced experience with Kubernetes and Amazon EKS, including workload deployment, cluster and application troubleshooting, observability, scaling, upgrades, access, and reliability.
  • Hands-on experience operating Argo Workflows or a comparable orchestration platform supporting production data workloads.
  • Strong experience with Infrastructure as Code, preferably Terraform, and with source-controlled configuration, CI/CD, release automation, and rollback practices.
  • Strong experience designing and operating observability, logging, monitoring, alerting, and incident-routing solutions for distributed production platforms.
  • Demonstrated ability to lead major incidents and cross-functional troubleshooting across infrastructure, applications, data pipelines, and analytics layers.
  • Experience operating production platforms with defined service levels, escalation paths, runbooks, change controls, release processes, and on-call responsibilities.
  • Strong working knowledge of modern data platforms, including Snowflake, S3-based data lakes, SQL, dbt, managed ingestion tools such as Fivetran or HVR, and custom data pipelines.
  • Understanding of data quality, freshness, lineage, schema evolution, pipeline dependencies, backfills, and recovery procedures.
  • Proficiency in Python, Shell, Bash, or comparable languages for automation and operational tooling.
  • Ability to make well-reasoned architecture and operational decisions, communicate tradeoffs, estimate work, identify risks, and guide teams through change.
  • Proven experience mentoring engineers, reviewing technical work, delegating ownership, and improving engineering processes across a distributed team.
  • Strong stakeholder-management and communication skills, including the ability to collect requirements, explain technical risks, and present recommendations to client leaders.
  • Strong written and verbal English communication skills.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed senior escalation and on-call rotation.

Responsibilities

  • Own the technical direction and end-to-end reliability of the managed Data Platform across AWS, Amazon EKS, Kubernetes, Argo Workflows, Snowflake, S3, Fivetran and custom ingestion pipelines, and Tableau dependencies.
  • Lead the technical transition into managed services, including platform discovery, dependency mapping, risk identification, knowledge transfer, shadowing, reverse shadowing, and readiness validation by platform tower.
  • Establish and evolve cloud architecture principles, Infrastructure as Code standards, CI/CD and release practices, observability patterns, operational controls, and platform engineering priorities.
  • Provide technical leadership for AWS infrastructure, Kubernetes and EKS operations, Argo Workflows, deployment automation, secrets and access management, platform capacity, and production reliability.
  • Translate business, service, security, and platform needs into actionable technical requirements, implementation plans, and a prioritized continuous-improvement backlog.
  • Guide cross-platform decisions involving Snowflake, dbt, Fivetran, custom ingestion, S3 data-lake operations, workflow orchestration, data quality, and downstream analytics dependencies.
  • Lead complex cross-platform troubleshooting and L3 escalation, coordinating engineers when incidents span infrastructure, ingestion, orchestration, Snowflake, and Tableau.
  • Direct technical response during major incidents, including impact assessment, stakeholder communication, recovery strategy, root-cause analysis, post-incident review, and preventive actions.
  • Ensure that changes and releases include appropriate technical review, dependency analysis, validation, rollback planning, traceability, and coordination across affected platform components.
  • Define and monitor reliability indicators such as availability, ingestion success, data freshness, workflow reliability, deployment outcomes, alert quality, mean time to restore service, and recurring incident patterns.
  • Lead improvements in automation, observability, auto-remediation, platform resilience, deployment safety, performance, security, capacity management, and cloud cost efficiency.
  • Partner with Security, Governance, Analytics, IT, and platform stakeholders to ensure least-privilege access, secrets management, auditability, controlled changes, and appropriate handling of regulated data.
  • Review technical designs and high-impact changes, challenge assumptions, document tradeoffs, and ensure solutions remain secure, scalable, and supportable within the managed-service operating model.
  • Mentor Senior and Middle-level engineers, delegate ownership effectively, improve team practices, and build consistent technical capability across the distributed team.
  • Maintain alignment between the LatAm and India coverage windows through clear ownership, escalation paths, operating procedures, and structured handoffs.
  • Participate in the senior escalation and on-call model for critical incidents outside staffed service hours.

Skills

AWS production
Kubernetes
Argo Workflows
Terraform
Python
CI/CD
Observability
Incident management

Tools

Snowflake
S3 data lake
dbt
Fivetran
HVR
Splunk
PagerDuty
Opsgenie

Job description

AgileEngine, LLC. is seeking a Cloud Platform Technical Lead in Puebla, Mexico to own the reliability and orchestration of an enterprise data platform in a regulated healthcare environment.

You will lead transition to managed services, define IaC and CI/CD standards, and oversee cross-functional incident response. The role requires deep AWS, Kubernetes/EKS, Argo Workflows, Terraform, Snowflake, S3 data lakes, and data pipeline experience, with strong English communication and LatAm hours

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud/DevOps Engineer - Remote AWS, Kubernetes & Argo
Senior Cloud/DevOps Engineer - Remote AWS, Kubernetes & Argo

AgileEngine, LLC. • Región Centro

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive pay
Exciting projects
+1
Senior Cloud Platform Lead - AWS, Kubernetes, Data Platform
Senior Cloud Platform Lead - AWS, Kubernetes, Data Platform

AgileEngine, LLC. • León

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud & DevOps Engineer — AWS, Kubernetes, Argo
Senior Cloud & DevOps Engineer — AWS, Kubernetes, Argo

AgileEngine, LLC. • Monterrey

On-site
MXN 1,200,000 - 2,000,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Cloud Platform Technical Lead - Reliability & Data Platform
Cloud Platform Technical Lead - Reliability & Data Platform

AgileEngine, LLC. • Santiago de Querétaro

On-site
MXN 2,172,000 - 3,258,000
Professional growth
USD-based pay
Exciting projects
+1
Remote Cloud Platform Lead (AWS/K8s)
Remote Cloud Platform Lead (AWS/K8s)

AgileEngine, LLC. • Rosarito

On-site
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation
Exciting projects
+1
Cloud Platform Tech Lead - Remote & High-Impact Reliability
Cloud Platform Tech Lead - Remote & High-Impact Reliability

AgileEngine, LLC. • Región Centro

Remote
MXN 2,534,000 - 3,620,000
Professional growth
USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer: AWS, Kubernetes, Terraform
Senior Cloud/DevOps Engineer: AWS, Kubernetes, Terraform

AgileEngine, LLC. • Puebla de Zaragoza

On-site
MXN 900,000 - 1,300,000
Senior Cloud DevOps Engineer (AWS, Kubernetes, Argo)
Senior Cloud DevOps Engineer (AWS, Kubernetes, Argo)

Agileengine • Xico

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive compensation: USD-based
Exciting projects
+1
Senior Cloud/DevOps Engineer — AWS, Kubernetes & Argo
Senior Cloud/DevOps Engineer — AWS, Kubernetes & Argo

AgileEngine, LLC. • Santiago de Querétaro

Hybrid
MXN 2,172,000 - 3,258,000
Professional growth
Competitive pay (USD)
Exciting projects
+1
Cloud Platform Lead: Reliability & Orchestration
Cloud Platform Lead: Reliability & Orchestration

AgileEngine, LLC. • Ciudad de México

Hybrid
MXN 1,200,000 - 1,900,000
Professional growth
Competitive compensation
Exciting projects
+1