Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC.

São Bernardo do Campo

Hybrid

BRL 602,000 - 903,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Professional growth
Competitive USD-based compensation
Exciting projects
Flextime

Job summary

AgileEngine is seeking a Senior Cloud/DevOps Engineer to operate and improve cloud infrastructure behind an enterprise data platform in a regulated healthcare environment. You will own the Cloud/DevOps service tower for the LatAm coverage window, manage AWS infrastructure, and lead incident response and observability initiatives.

The role requires 5+ years in Cloud Engineering, DevOps or SRE, strong English communication, and hands-on experience with AWS EKS, Kubernetes, Argo Workflows,

Qualifications

  • 5+ years in Cloud Engineering, DevOps or SRE.
  • Strong hands-on experience with AWS infrastructure in production.
  • Advanced experience with Kubernetes and Amazon EKS.
  • Hands-on experience with Argo Workflows or similar.
  • Infrastructure as Code with Terraform and Git-based delivery.
  • Experience building and supporting CI/CD pipelines and releases.
  • Strong monitoring/logging/alerting with tools like Splunk/PagerDuty/Opsgenie.
  • Root-cause analysis and incident resolution skills.
  • Proficiency in Python, Shell or Bash for automation.
  • Ability to mentor engineers and collaborate with data, security and governance teams.

Responsibilities

  • Own the Cloud/DevOps service tower for LatAm coverage window.
  • Operate and improve AWS infrastructure including EKS, S3, EventBridge, SQS, API Gateway, Lambda.
  • Administer Kubernetes workloads and Argo Workflows, including deployment and troubleshooting.
  • Define and improve IaC standards using Terraform and Git-based practices.
  • Lead observability improvements across infra and data workloads.
  • Improve alert routing and escalation across Splunk, Opsgenie, PagerDuty, Teams.
  • Lead incident response and post-incident reviews with stakeholders.
  • Design reliability improvements and automate repetitive tasks.
  • Track service metrics like availability, SLA, and MTTR.
  • Ensure change-control, access control, secrets management for HIPAA/GDPR/FDA environments.
  • Create runbooks and knowledge transfer materials.
  • Mentor mid-level engineers and promote best practices.

Skills

Cloud Engineering
DevOps
Kubernetes
Argo Workflows
Terraform
CI/CD pipelines
Incident response
Automation scripting
Leadership

Tools

AWS EKS
Argo Workflows
Splunk
PagerDuty
Opsgenie
Terraform
Jira

Job description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a Senior Cloud/DevOps Engineer to operate and improve the cloud infrastructure and reliability layers behind an enterprise data platform in a regulated healthcare environment.

The mandatory requirements are 5+ years of experience in Cloud Engineering, DevOps, or Site Reliability Engineering, advanced hands-on experience with AWS EKS and Kubernetes, experience administering and troubleshooting Argo Workflows, and strong English communication skills.

MUST HAVES
  • 5+ years of professional experience in Cloud Engineering, DevOps or Site Reliability Engineering.
  • Strong hands‑on experience operating AWS infrastructure in production environments.
  • Advanced experience with Kubernetes and Amazon EKS, including workload operations, troubleshooting, access, observability, capacity, and reliability.
  • Hands‑on experience administering and troubleshooting Argo Workflows or comparable workflow orchestration platforms.
  • Strong Infrastructure as Code experience with Terraform and source‑controlled infrastructure practices.
  • Experience building, hardening, and supporting CI/CD pipelines and production release processes.
  • Strong experience with monitoring, logging, alerting, and incident‑routing tools such as Splunk, PagerDuty, Opsgenie, or comparable platforms.
  • Demonstrated ability to lead complex incident resolution, perform root‑cause analysis, and translate findings into preventive improvements.
  • Proficiency in automation and scripting using Python, Shell, Bash, or similar languages.
  • Ability to make well‑reasoned technical decisions, identify tradeoffs, estimate work, and drive improvements across a complex platform.
  • Experience mentoring engineers and collaborating effectively with Data Engineering, Security, Governance, Analytics, and business stakeholders.
  • Strong written and verbal English communication skills, with the ability to work directly with client stakeholders.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed on‑call rotation.
NICE TO HAVES
  • Experience supporting data‑platform infrastructure involving Snowflake, dbt, Fivetran, HVR, Tableau Cloud, or custom ingestion pipelines.
  • Familiarity with data‑specific observability platforms such as SYNQ.
  • Experience modernizing or migrating legacy orchestration and ingestion solutions such as Boomi or AWS Data Pipeline.
  • Experience with service‑management and change‑control tools such as Freshservice and Jira.
  • Experience operating in healthcare, life sciences, financial services, or another regulated environment.
  • Familiarity with HIPAA, GDPR, FDA‑related controls, least‑privilege access, separation of duties, and audit‑ready operational practices.
WHAT YOU WILL DO
  • Provide senior technical ownership for the Cloud / DevOps service tower during the LatAm coverage window, including day‑to‑day operations, complex troubleshooting, and L2/L3 escalation.
  • Operate, maintain, and improve AWS infrastructure supporting the Data Platform, including Amazon EKS, S3, EventBridge, SQS, API Gateway, Lambda, and related services.
  • Administer Kubernetes‑hosted workloads and Argo Workflows, including deployment, scheduling, monitoring, troubleshooting, capacity management, resiliency, and recovery.
  • Define and improve standards for Infrastructure as Code, configuration management, CI/CD, release execution, rollback, and environment consistency, primarily using Terraform and Git‑based delivery practices.
  • Lead the consolidation and improvement of observability across infrastructure and data workloads, linking alerts to operational evidence from Argo, dbt, Snowflake, and supporting runbooks.
  • Improve alert routing and escalation workflows across tools such as Splunk, Opsgenie, PagerDuty, Microsoft Teams, and data‑specific observability platforms.
  • Lead or support major incident response, root‑cause analysis, post‑incident reviews, and corrective actions, with clear communication to technical and service stakeholders.
  • Design and implement reliability improvements such as selective auto‑remediation, dependency‑aware alert correlation, impact analysis, and automation of repetitive operational work.
  • Track and contribute to service metrics including availability, SLA compliance, alert volumes, workflow reliability, deployment outcomes, and mean time to restore service.
  • Apply disciplined change‑management, access‑control, secrets‑management, auditability, and documentation practices appropriate for a HIPAA-, GDPR-, and FDA‑regulated environment.
  • Create and maintain runbooks, operating procedures, architecture context, recovery procedures, and knowledge‑transfer materials.
  • Mentor Middle‑level engineers, review technical work, improve team practices, and promote consistent execution across the distributed team.
  • Participate in the Cloud / DevOps on‑call rotation for critical incidents outside staffed service hours.
PERKS AND BENEFITS
  • Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
  • Competitive compensation: USD‑based pay with education, fitness, and team activity budgets.
  • Exciting projects: Modern solutions with Fortune 500 and top product companies.
  • Flextime: Flexible schedule with remote and office options.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • São Paulo

Hybrid
BRL 260,000 - 380,000
Professional growth
Competitive USD-based compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Salvador

Hybrid
BRL 602,000 - 903,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Florianópolis

Hybrid
BRL 602,000 - 903,000
Professional growth
Competitive compensation USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Brasília

Remote
BRL 803,000 - 1,104,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Campinas

Hybrid
BRL 602,000 - 903,000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Rio de Janeiro

Hybrid
BRL 602,000 - 903,000
Professional growth
Competitive USD-based compensation
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Curitiba

Hybrid
BRL 502,000 - 753,000
Professional growth
Competitive USD pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Sorocaba

Hybrid
BRL 180,000 - 320,000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Senior Cloud/DevOps Engineer ID92208
Senior Cloud/DevOps Engineer ID92208

AgileEngine, LLC. • Belo Horizonte

Hybrid
BRL 602,000 - 803,000
Professional growth
Competitive USD pay
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine • Sorocaba

Hybrid
BRL 750,000 - 1,051,000
Professional growth
Competitive compensation
Exciting projects
+1