Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC.

Bogotá ciudad

On-site

COP 388,387,000 - 582,581,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Mentorship program
TechTalks & growth roadmap
USD-based compensation
Flexible schedule (remote/office)

Job summary

AgileEngine, LLC. seeks a Senior Cloud/DevOps Engineer to run and improve the cloud infrastructure behind an enterprise data platform in a regulated healthcare environment.

This role requires 5+ years in Cloud Eng/DevOps/SRE with hands-on AWS EKS, Kubernetes, Argo Workflows, and Terraform expertise, plus strong English communication. You will lead incident response, automate resilience, and collaborate with data, security, governance, and analytics teams.

Qualifications

  • 5+ years of experience in Cloud Eng, DevOps or SRE.
  • Hands-on AWS production experience (EKS).
  • Advanced Kubernetes and EKS workload operations.
  • Hands-on Argo Workflows administration & troubleshooting.
  • Terraform IaC and Git-based delivery practices.
  • CI/CD pipeline design and production releases.
  • Monitoring/logging with incident routing tools (Splunk/ PagerDuty).
  • Scripting in Python, Shell, or Bash.
  • Strong English communication; client-facing ability.
  • Ability to mentor engineers and collaborate with cross-functional teams.

Responsibilities

  • Provide senior technical ownership for the Cloud/DevOps service tower during LatAm hours.
  • Operate and improve AWS infrastructure, including EKS, S3, EventBridge, API Gateway, Lambda.
  • Administer Kubernetes workloads and Argo Workflows; ensure deployment and recovery.
  • Define and improve standards for IaC, configuration, CI/CD, and environment consistency.
  • Lead observability improvements; link alerts to runbooks and data platforms.
  • Improve alert routing across Splunk, Opsgenie, PagerDuty, and Teams.
  • Lead major incident response, root-cause analysis, and post-incident reviews.
  • Design reliability improvements: auto-remediation and informed tradeoffs.
  • Track service metrics: availability, SLA, alert volumes, and MTTR.
  • Ensure HIPAA/GDPR/FDA controls, access management, and auditability.
  • Create runbooks and knowledge transfer materials; mentor engineers.

Skills

5+ years Cloud Eng/DevOps/SRE
AWS & EKS experience
Kubernetes expertise
Argo Workflows
Terraform / IaC
CI/CD pipelines
Monitoring & incident response
Python / Shell scripting
Leadership & mentoring
English communication

Tools

AWS EKS
Kubernetes
Argo Workflows
Terraform
Git-based delivery
Splunk / PagerDuty / Opsgenie
CI/CD tooling
Python scripting

Job description


  • Country Colombia

  • Job Type Full time


Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.


WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!


ABOUT THE ROLE

We are looking for a Senior Cloud/DevOps Engineer to operate and improve the cloud infrastructure and reliability layers behind an enterprise data platform in a regulated healthcare environment.


The mandatory requirements are 5+ years of experience in Cloud Engineering, DevOps, or Site Reliability Engineering, advanced hands-on experience with AWS EKS and Kubernetes, experience administering and troubleshooting Argo Workflows, and strong English communication skills.


MUST HAVES


  • - 5+ years of professional experience in Cloud Engineering, DevOps or Site Reliability Engineering.

  • - Strong hands-on experience operating AWS infrastructure in production environments.

  • - Advanced experience with Kubernetes and Amazon EKS, including workload operations, troubleshooting, access, observability, capacity, and reliability.

  • - Hands-on experience administering and troubleshooting Argo Workflows or comparable workflow orchestration platforms.

  • - Strong Infrastructure as Code experience with Terraform and source-controlled infrastructure practices.

  • - Experience building, hardening, and supporting CI/CD pipelines and production release processes.

  • - Strong experience with monitoring, logging, alerting, and incident-routing tools such as Splunk, PagerDuty, Opsgenie, or comparable platforms.

  • - Demonstrated ability to lead complex incident resolution, perform root‑cause analysis, and translate findings into preventive improvements.

  • - Proficiency in automation and scripting using Python, Shell, Bash, or similar languages.

  • - Ability to make well-reasoned technical decisions, identify tradeoffs, estimate work, and drive improvements across a complex platform.

  • - Experience mentoring engineers and collaborating effectively with Data Engineering, Security, Governance, Analytics, and business stakeholders.

  • - Strong written and verbal English communication skills, with the ability to work directly with client stakeholders.

  • - Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed on‑call rotation.


NICE TO HAVES


  • - Experience supporting data-platform infrastructure involving Snowflake, dbt, Fivetran, HVR, Tableau Cloud, or custom ingestion pipelines.

  • - Familiarity with data-specific observability platforms such as SYNQ.

  • - Experience modernizing or migrating legacy orchestration and ingestion solutions such as Boomi or AWS Data Pipeline.

  • - Experience with service‑management and change‑control tools such as Freshservice and Jira.

  • - Experience operating in healthcare, life sciences, financial services, or another regulated environment.

  • - Familiarity with HIPAA, GDPR, FDA‑related controls, least‑privilege access, separation of duties, and audit‑ready operational practices.


WHAT YOU WILL DO


  • - Provide senior technical ownership for the Cloud / DevOps service tower during the LatAm coverage window, including day‑to‑day operations, complex troubleshooting, and L2/L3 escalation.

  • - Operate, maintain, and improve AWS infrastructure supporting the Data Platform, including Amazon EKS, S3, EventBridge, SQS, API Gateway, Lambda, and related services.

  • - Administer Kubernetes‑hosted workloads and Argo Workflows, including deployment, scheduling, monitoring, troubleshooting, capacity management, resiliency, and recovery.

  • - Define and improve standards for Infrastructure as Code, configuration management, CI/CD, release execution, rollback, and environment consistency, primarily using Terraform and Git‑based delivery practices.

  • - Lead the consolidation and improvement of observability across infrastructure and data workloads, linking alerts to operational evidence from Argo, dbt, Snowflake, and supporting runbooks.

  • - Improve alert routing and escalation workflows across tools such as Splunk, Opsgenie, PagerDuty, Microsoft Teams, and data‑specific observability platforms.

  • - Lead or support major incident response, root‑cause analysis, post‑incident reviews, and corrective actions, with clear communication to technical and service stakeholders.

  • - Design and implement reliability improvements such as selective auto‑remediation, dependency‑aware alert correlation, impact analysis, and automation of repetitive operational work.

  • - Track and contribute to service metrics including availability, SLA compliance, alert volumes, workflow reliability, deployment outcomes, and mean time to restore service.

  • - Apply disciplined change‑management, access‑control, secrets‑management, auditability, and documentation practices appropriate for a HIPAA-, GDPR-, and FDA‑regulated environment.

  • - Create and maintain runbooks, operating procedures, architecture context, recovery procedures, and knowledge‑transfer materials.

  • - Mentor Middle-level engineers, review technical work, improve team practices, and promote consistent execution across the distributed team.

  • - Participate in the Cloud / DevOps on‑call rotation for critical incidents outside staffed service hours.


PERKS AND BENEFITS


  • - Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.

  • - Competitive compensation: USD-based pay with education, fitness, and team activity budgets.

  • - Exciting projects: Modern solutions with Fortune 500 and top product companies.

  • - Flextime: Flexible schedule with remote and office options.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Capital

Hybrid
COP 388,387,000 - 582,581,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Metropolitana

Hybrid
COP 388,387,000 - 582,581,000
Professional growth
USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Perímetro Urbano Manizales

Hybrid
COP 323,656,000 - 534,032,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Medellín

Hybrid
COP 291,290,000 - 517,850,000
Professional growth
Competitive compensation USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Sur

Remote
COP 388,387,000 - 517,850,000
Professional growth
Competitive compensation: USD-based
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Cartagena de Indias

On-site
COP 291,290,000 - 453,118,000
Professional growth
USD-based pay with budgets
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Hybrid
COP 291,290,000 - 453,118,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Pereira

Hybrid
COP 388,387,000 - 517,850,000
Professional growth
Competitive USD pay
Exciting projects
+1
Senior Cloud/Devops Engineer Id92208
Senior Cloud/Devops Engineer Id92208

Agileengine • Sucre

Remote
COP 388,387,000 - 582,581,000
Professional growth
Competitive compensation: USD-basedpay
Exciting projects
+1
Senior Cloud/Devops Engineer Id92208
Senior Cloud/Devops Engineer Id92208

Agileengine • Risaralda

Remote
COP 323,656,000 - 485,484,000
Professional growth
Competitive pay
Exciting projects
+1