Cloud Platform Technical Lead Id92209

Agileengine

San Miguel de Tucumán

Híbrido

ARS 228.464.000 - 319.849.000

Jornada completa

Hace 6 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Professional growth
Competitive compensation
Exciting projects
Flextime

Descripción de la vacante

AgileEngine is seeking a Cloud Platform Technical Lead to own the reliability and orchestration of an enterprise data platform in a regulated healthcare environment. You will lead AWS/EKS operations, Argo workflows, IaC with Terraform, and platform observability, while mentoring engineers across a distributed team.

You will drive incident response, architecture decisions, and cross-functional collaboration with security, analytics, and IT stakeholders, supporting a global client base with

Formación

  • 6+ years in Cloud/Platform Engineering, DevOps or SRE.
  • Deep AWS production environment expertise with security, networking and troubleshooting.
  • Kubernetes and Amazon EKS experience with deployments, observability, and reliability.
  • Hands-on with Argo Workflows or similar orchestration for data workloads.
  • IaC experience, Terraform preferred, with CI/CD and release automation.
  • Design and operate observability, logging, monitoring, alerting, and incident routing.

Responsabilidades

  • Own reliability and end-to-end direction of the data platform on AWS, EKS, Snowflake, S3, and related tools.
  • Lead transition to managed services, with discovery, risk mapping, and knowledge transfer.
  • Define cloud architecture principles, IaC standards, CI/CD, and release practices.
  • Provide leadership for AWS infra, Kubernetes, Argo, secrets management, capacity, and on-call reliability.
  • Translate business and security needs into technical requirements and backlogs.
  • Mentor engineers and drive improvements across distributed teams.

Conocimientos

Cloud platform engineering
AWS expertise
Kubernetes & EKS
Argo Workflows
Infrastructure as Code (Terraform)
Observability & monitoring
Incident management
Team leadership & mentoring
Cross-functional collaboration
English communication

Herramientas

Terraform
Python
Shell/Bash
Argo Workflows

Descripción del empleo

Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a Cloud Platform Technical Lead to own the reliability and orchestration of an enterprise data platform in a regulated healthcare environment.

MUST HAVES
  • 6+ years of professional experience in Cloud Engineering, Platform Engineering, DevOps or Site Reliability Engineering
  • Deep hands-on expertise with AWS production environments, including cloud architecture, security, networking fundamentals, access management, monitoring, capacity, and operational troubleshooting.
  • Advanced experience with Kubernetes and Amazon EKS, including workload deployment, cluster and application troubleshooting, observability, scaling, upgrades, access, and reliability.
  • Hands-on experience operating Argo Workflows or a comparable orchestration platform supporting production data workloads.
  • Strong experience with Infrastructure as Code, preferably Terraform, and with source-controlled configuration, CI/CD, release automation, and rollback practices.
  • Strong experience designing and operating observability, logging, monitoring, alerting, and incident-routing solutions for distributed production platforms.
  • Demonstrated ability to lead major incidents and cross-functional troubleshooting across infrastructure, applications, data pipelines, and analytics layers.
  • Experience operating production platforms with defined service levels, escalation paths, runbooks, change controls, release processes, and on-call responsibilities.
  • Strong working knowledge of modern data platforms, including Snowflake, S3-based data lakes, SQL, dbt, managed ingestion tools such as Fivetran or HVR, and custom data pipelines.
  • Understanding of data quality, freshness, lineage, schema evolution, pipeline dependencies, backfills, and recovery procedures.
  • Proficiency in Python, Shell, Bash or comparable languages for automation and operational tooling.
  • Ability to make well-reasoned architecture and operational decisions, communicate tradeoffs, estimate work, identify risks, and guide teams through change.
  • Proven experience mentoring engineers, reviewing technical work, delegating ownership, and improving engineering processes across a distributed team.
  • Strong stakeholder-management and communication skills, including the ability to collect requirements, explain technical risks, and present recommendations to client leaders.
  • Strong written and verbal English communication skills.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed senior escalation and on-call rotation.
NICE TO HAVES
  • Experience with data observability platforms such as SYNQ and operational tooling such as Splunk, PagerDuty, Opsgenie, or comparable solutions.
  • Experience implementing dependency-aware alerting, selective auto-remediation, impact analysis, or other advanced reliability practices.
  • Experience leading a managed service, platform operations team, Site Reliability Engineering function, or follow-the-sun support model.
  • Experience in healthcare, financial services, or another regulated environment.
WHAT YOU WILL DO
  • Own the technical direction and end-to-end reliability of the managed Data Platform across AWS, Amazon EKS, Kubernetes, Argo Workflows, Snowflake, S3, Fivetran and custom ingestion pipelines, and Tableau dependencies.
  • Lead the technical transition into managed services, including platform discovery, dependency mapping, risk identification, knowledge transfer, shadowing, reverse shadowing, and readiness validation by platform tower.
  • Establish and evolve cloud architecture principles, Infrastructure as Code standards, CI/CD and release practices, observability patterns, operational controls, and platform engineering priorities.
  • Provide technical leadership for AWS infrastructure, Kubernetes and EKS operations, Argo Workflows, deployment automation, secrets and access management, platform capacity, and production reliability.
  • Translate business, service, security, and platform needs into actionable technical requirements, implementation plans, and a prioritized continuous-improvement backlog.
  • Guide cross-platform decisions involving Snowflake, dbt, Fivetran, custom ingestion, S3 data-lake operations, workflow orchestration, data quality, and downstream analytics dependencies.
  • Lead complex cross-platform troubleshooting and L3 escalation, coordinating engineers when incidents span infrastructure, ingestion, orchestration, Snowflake, and Tableau.
  • Direct technical response during major incidents, including impact assessment, stakeholder communication, recovery strategy, root-cause analysis, post-incident review, and preventive actions.
  • Ensure that changes and releases include appropriate technical review, dependency analysis, validation, rollback planning, traceability, and coordination across affected platform components.
  • Define and monitor reliability indicators such as availability, ingestion success, data freshness, workflow reliability, deployment outcomes, alert quality, mean time to restore service, and recurring incident patterns.
  • Lead improvements in automation, observability, auto-remediation, platform resilience, deployment safety, performance, security, capacity management, and cloud cost efficiency.
  • Partner with Security, Governance, Analytics, IT, and platform stakeholders to ensure least-privilege access, secrets management, auditability, controlled changes, and appropriate handling of regulated data.
  • Review technical designs and high-impact changes, challenge assumptions, document tradeoffs, and ensure solutions remain secure, scalable, and supportable within the managed-service operating model.
  • Mentor Senior and Middle-level engineers, delegate ownership effectively, improve team practices, and build consistent technical capability across the distributed team.
  • Maintain alignment between the LatAm and India coverage windows through clear ownership, escalation paths, operating procedures, and structured handoffs.
  • Participate in the senior escalation and on-call model for critical incidents outside staffed service hours.
PERKS AND BENEFITS
  • Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
  • Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
  • Exciting projects: Modern solutions with Fortune 500 and top product companies.
  • Flextime: Flexible schedule with remote and office options.
Requirements
  • 6+ years of professional experience in Cloud Engineering, Platform Engineering, DevOps or Site Reliability Engineering
  • Deep hands-on expertise with AWS production environments, including cloud architecture, security, networking fundamentals, access management, monitoring, capacity, and operational troubleshooting.
  • Advanced experience with Kubernetes and Amazon EKS, including workload deployment, cluster and application troubleshooting, observability, scaling, upgrades, access, and reliability.
  • Hands-on experience operating Argo Workflows or a comparable orchestration platform supporting production data workloads.
  • Strong experience with Infrastructure as Code, preferably Terraform, and with source-controlled configuration, CI/CD, release automation, and rollback practices.
  • Strong experience designing and operating observability, logging, monitoring, alerting, and incident-routing solutions for distributed production platforms.
  • Demonstrated ability to lead major incidents and cross-functional troubleshooting across infrastructure, applications, data pipelines, and analytics layers.
  • Experience operating production platforms with defined service levels, escalation paths, runbooks, change controls, release processes, and on-call responsibilities.
  • Strong working knowledge of modern data platforms, including Snowflake, S3-based data lakes, SQL, dbt, managed ingestion tools such as Fivetran or HVR, and custom data pipelines.
  • Understanding of data quality, freshness, lineage, schema evolution, pipeline dependencies, backfills, and recovery procedures.
  • Proficiency in Python, Shell, Bash, or comparable languages for automation and operational tooling.
  • Ability to make well-reasoned architecture and operational decisions, communicate tradeoffs, estimate work, identify risks, and guide teams through change.
  • Proven experience mentoring engineers, reviewing technical work, delegating ownership, and improving engineering processes across a distributed team.
  • Strong stakeholder-management and communication skills, including the ability to collect requirements, explain technical risks, and present recommendations to client leaders.
  • Strong written and verbal English communication skills.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed senior escalation and on-call rotation.
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Cloud Platform Technical Lead Id92209
Cloud Platform Technical Lead Id92209

Agileengine • San Carlos de Bariloche

A distancia
ARS 213.233.000 - 350.311.000
Professional growth
Competitive compensation (USD)
Exciting projects
+1
Tech Lead .Net: Líder Técnico De Microservicios (Remoto)
Tech Lead .Net: Líder Técnico De Microservicios (Remoto)

Skydropx - Frenet • Ciudad de Mendoza

A distancia
ARS 182.310.000 - 273.465.000
Professional growth
Competitive USD-based pay
Exciting projects
+1
Tech Lead .Net: Líder Técnico De Microservicios (Remoto)
Tech Lead .Net: Líder Técnico De Microservicios (Remoto)

Skydropx - Frenet • Rosario

A distancia
ARS 258.272.000 - 319.042.000
Professional growth
Competitive compensation
Exciting projects
+1
Senior Devsecops Engineer — Aws Cloud Platform & Sre
Senior Devsecops Engineer — Aws Cloud Platform & Sre

Talan • Ciudad de Mendoza

Presencial
ARS 212.695.000 - 273.465.000
Professional growth
Competitive compensation
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine, LLC. • San Carlos de Bariloche

Híbrido
ARS 182.771.000 - 274.156.000
Professional growth
Competitive compensation
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine, LLC. • Córdoba

Híbrido
ARS 213.233.000 - 304.618.000
Professional growth
Competitive compensation
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine, LLC. • Ciudad de Mendoza

Presencial
ARS 213.233.000 - 289.387.000
Professional growth
Competitive compensation
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine, LLC. • Rosario

Híbrido
ARS 228.464.000 - 289.387.000
Professional growth
Competitive compensation
Exciting projects
+1
Tech Lead (.Net)
Tech Lead (.Net)

Skydropx - Frenet • Buenos Aires

Híbrido
ARS 213.233.000 - 289.387.000
Professional growth
USD-based pay
Exciting projects
+1
Cloud Platform Technical Lead ID92209
Cloud Platform Technical Lead ID92209

AgileEngine • Ciudad de Mendoza

Híbrido
ARS 212.899.000 - 304.141.000
Professional growth
Competitive compensation
Exciting projects
+1