Senior Cloud & DevOps Engineer — AWS EKS, Kubernetes & Argo

AgileEngine

United States

Remote

MXN 900,000 - 1,300,000

Full time

16 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

100% remote work
Annual learning budget
Well-being programs

Job summary

AgileEngine is seeking a Senior Cloud/DevOps Engineer to own and improve the cloud infrastructure supporting a regulated healthcare data platform. You will manage AWS EKS, Kubernetes workloads, Argo Workflows, and IaC with Terraform, while enhancing observability, incident response, and release processes.

The role requires 5+ years in Cloud/DevOps or SRE, strong English communication, and availability for LatAm hours. 100% remote options may apply across the region.

Qualifications

  • 5+ years of professional experience in Cloud Engineering, DevOps or Site Reliability Engineering.
  • Strong hands-on experience operating AWS infrastructure in production environments.
  • Advanced experience with Kubernetes and Amazon EKS, including workload operations, troubleshooting, access, observability, capacity, and reliability.
  • Hands-on experience administering and troubleshooting Argo Workflows or comparable workflow orchestration platforms.
  • Strong Infrastructure as Code experience with Terraform and source-controlled infrastructure practices.
  • Experience building, hardening, and supporting CI/CD pipelines and production release processes.
  • Strong experience with monitoring, logging, alerting, and incident-routing tools such as Splunk, PagerDuty, Opsgenie, or comparable platforms.
  • Demonstrated ability to lead complex incident resolution, perform root-cause analysis, and translate findings into preventive improvements.
  • Proficiency in automation and scripting using Python, Shell, Bash, or similar languages.
  • Ability to make well-reasoned technical decisions, identify tradeoffs, estimate work, and drive improvements across a complex platform.
  • Experience mentoring engineers and collaborating effectively with Data Engineering, Security, Governance, Analytics, and business stakeholders.
  • Strong written and verbal English communication skills.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed on-call rotation.

Responsibilities

  • Provide senior technical ownership for the Cloud / DevOps service tower during the LatAm coverage window, including day-to-day operations, complex troubleshooting, and L2/L3 escalation.
  • Operate, maintain, and improve AWS infrastructure supporting the Data Platform, including Amazon EKS, S3, EventBridge, SQS, API Gateway, Lambda, and related services.
  • Administer Kubernetes-hosted workloads and Argo Workflows, including deployment, scheduling, monitoring, troubleshooting, capacity management, resiliency, and recovery.
  • Define and improve standards for Infrastructure as Code, configuration management, CI/CD, release execution, rollback, and environment consistency, primarily using Terraform and Git-based delivery practices.
  • Lead the consolidation and improvement of observability across infrastructure and data workloads, linking alerts to operational evidence from Argo, dbt, Snowflake, and supporting runbooks.
  • Improve alert routing and escalation workflows across tools such as Splunk, Opsgenie, PagerDuty, Microsoft Teams, and data-specific observability platforms.
  • Lead or support major incident response, root-cause analysis, post-incident reviews, and corrective actions, with clear communication to technical and service stakeholders.
  • Design and implement reliability improvements such as selective auto-remediation, dependency-aware alert correlation, impact analysis, and automation of repetitive operational work.
  • Track and contribute to service metrics including availability, SLA compliance, alert volumes, workflow reliability, deployment outcomes, and mean time to restore service.
  • Apply disciplined change-management, access-control, secrets-management, auditability, and documentation practices appropriate for a HIPAA-, GDPR-, and FDA-regulated environment.
  • Create and maintain runbooks, operating procedures, architecture context, recovery procedures, and knowledge-transfer materials.
  • Mentor Middle-level engineers, review technical work, improve team practices, and promote consistent execution across the distributed team.
  • Participate in the Cloud / DevOps on-call rotation for critical incidents outside staffed service hours.

Skills

English communication
Mentoring
Leadership
Problem solving
On-call coordination

Tools

AWS
Kubernetes
Argo Workflows
Terraform
Git
Splunk
PagerDuty
EventBridge

Job description

AgileEngine is seeking a Senior Cloud/DevOps Engineer to own and improve the cloud infrastructure supporting a regulated healthcare data platform. You will manage AWS EKS, Kubernetes workloads, Argo Workflows, and IaC with Terraform, while enhancing observability, incident response, and release processes.

The role requires 5+ years in Cloud/DevOps or SRE, strong English communication, and availability for LatAm hours. 100% remote options may apply across the region.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud & Kubernetes DevOps Engineer (Remote)
Senior Cloud & Kubernetes DevOps Engineer (Remote)

PAR Technology • Illinois

On-site
USD 155,000 - 170,000
Senior Cloud Platform Engineer - Kubernetes, GCP, GitOps
Senior Cloud Platform Engineer - Kubernetes, GCP, GitOps

United States Digital Space LLC • United States

Remote
USD 140,000 - 190,000
Lead DevOps Engineer - AWS, Kubernetes, GitOps (Remote)
Lead DevOps Engineer - AWS, Kubernetes, GitOps (Remote)

PAR Technology • North Carolina

On-site
USD 155,000 - 170,000
Senior Cloud SRE - Kubernetes, Terraform, Remote
Senior Cloud SRE - Kubernetes, Terraform, Remote

ascendingdc • United States

Remote
USD 140,000 - 190,000
Senior DevOps Systems Engineer - AWS, Kubernetes, CI/CD
Senior DevOps Systems Engineer - AWS, Kubernetes, CI/CD

Avid Technology Professionals • Maryland

On-site
USD 120,000 - 150,000
Remote Senior DevOps Engineer - AWS, Kubernetes, GitOps
Remote Senior DevOps Engineer - AWS, Kubernetes, GitOps

PAR Technology • Georgia

On-site
USD 155,000 - 170,000
Remote Site Reliability Engineer – Kubernetes & CloudOps
Remote Site Reliability Engineer – Kubernetes & CloudOps

Sporty Group • United States

Remote
USD 140,000 - 190,000
Remote-first company
Quarterly performance bonuses
28 days paid annual leave
+4
Remote Senior Platform Engineer (AWS, Kubernetes, CI/CD)
Remote Senior Platform Engineer (AWS, Kubernetes, CI/CD)

Ledgebrook Insurance, LLC. • United States

Remote
USD 150,000 - 200,000
Unlimited PTO
Remote work
Health benefits
+4
Remote Senior DevOps Engineer – AWS, Kubernetes & GitOps
Remote Senior DevOps Engineer – AWS, Kubernetes & GitOps

PAR Technology • Minnesota

On-site
USD 155,000 - 170,000
Remote Senior DevOps Engineer: AWS, Kubernetes, GitOps
Remote Senior DevOps Engineer: AWS, Kubernetes, GitOps

PAR Technology • New York (NY)

On-site
USD 155,000 - 170,000