Cloud Platform Technical Lead

AgileEngine, LLC.

Brasil

Teletrabalho

BRL 350 000 - 520 000

Tempo integral

há 47 horas
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Não envies um currículo genérico — gera um currículo e uma carta de apresentação adaptados a esta função específica.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Growth without limits
Competitive compensation
Remote work with flexible hours
Modern projects with global teams
Collaborative culture
Well-being programs

Resumo da oferta

AgileEngine, LLC. in Brazil seeks a Cloud Platform Technical Lead to own the reliability and orchestration of an enterprise data platform in a regulated healthcare setting.

You will steer end-to-end platform reliability across AWS EKS, Argo Workflows, Snowflake, and data pipelines while guiding cross-functional teams. You will design cloud architecture, IaC patterns, CI/CD practices, and incident response, mentoring engineers and collaborating with Security, Governance, Analytics, IT, and client

Qualificações

  • 6+ years of professional experience in Cloud/Platform Engineering, DevOps or SRE.
  • Hands-on AWS production environments including cloud architecture, security, networking, access management, monitoring, capacity, and troubleshooting.
  • Advanced experience with Kubernetes and Amazon EKS, including workload deployment, observability, scaling, upgrades, access, and reliability.
  • Hands-on experience operating Argo Workflows or comparable orchestration for production data workloads.
  • Strong IaC experience, preferably Terraform, and with CI/CD, release automation, and rollback practices.
  • Experience designing and operating observability, logging, monitoring, alerting, and incident routing for distributed platforms.
  • Proven ability to lead major incidents and cross-functional troubleshooting across infrastructure, apps, data pipelines, and analytics.
  • Knowledge of data platforms like Snowflake, S3 data lakes, SQL, dbt; data quality, lineage, backfills, and recovery.
  • Proficiency in Python, Shell, Bash for automation.
  • Ability to communicate tradeoffs, estimate work, and guide teams through change.
  • Mentor engineers and improve engineering processes across distributed teams.
  • Strong stakeholder management and English communication skills.
  • Availability to work within LatAm service window (9:00–18:00 ET) and participate in on-call rotation.

Responsabilidades

  • Own the technical direction and end-to-end reliability of the managed Data Platform across AWS, EKS, Kubernetes, Argo, Snowflake, S3, Fivetran and Tableau dependencies.
  • Lead the technical transition into managed services, including platform discovery, mapping, risk identification, knowledge transfer, and readiness validation.
  • Establish and evolve cloud architecture principles, IaC standards, CI/CD, observability patterns, and platform priorities.
  • Provide technical leadership for AWS infrastructure, Kubernetes/EKS operations, Argo Workflows, deployment automation, secrets/access management, capacity, and reliability.
  • Translate business and security needs into actionable technical requirements and backlogs.
  • Guide cross-platform decisions involving Snowflake, dbt, Fivetran, ingestion pipelines and analytics dependencies.
  • Lead cross-platform troubleshooting and L3 escalation during incidents.
  • Direct major incident response with review, recovery, and preventive actions.
  • Ensure changes include dependency analysis, validation, rollback planning, and traceability.
  • Define and monitor reliability indicators such as availability, data freshness, workflow reliability, and MTTR.
  • Lead automation, observability, auto-remediation, and cloud cost efficiency improvements.
  • Partner with Security, Governance, Analytics, IT to ensure least-privilege access and auditability.
  • Review designs and ensure secure, scalable solutions within managed-service model.
  • Mentor senior engineers and build distributed team capability.
  • Maintain alignment across LatAm and India coverage windows.

Conhecimentos

AWS
Kubernetes
Argo Workflows
Snowflake
SQL
dbt
Python
Shell/Bash
Terraform
Observability

Ferramentas

Terraform
Argo Workflows
Kubernetes
EKS
AWS

Descrição da oferta de emprego

About the role

We are looking for a Cloud Platform Technical Lead to own the reliability and orchestration of an enterprise data platform in a regulated healthcare environment.

The mandatory requirements are 6+ years of experience in Cloud Engineering, Platform Engineering, DevOps, or Site Reliability Engineering, hands‑on experience with AWS EKS and data orchestration using Argo Workflows, strong working knowledge of modern data platforms including Snowflake, S3-based data lakes, SQL, and dbt, and strong English communication skills.

Must haves
  • 6+ years of professional experience in Cloud Engineering, Platform Engineering, DevOps or Site Reliability Engineering
  • Deep hands‑on expertise with AWS production environments, including cloud architecture, security, networking fundamentals, access management, monitoring, capacity, and operational troubleshooting.
  • Advanced experience with Kubernetes and Amazon EKS, including workload deployment, cluster and application troubleshooting, observability, scaling, upgrades, access, and reliability.
  • Hands‑on experience operating Argo Workflows or a comparable orchestration platform supporting production data workloads.
  • Strong experience with Infrastructure as Code, preferably Terraform, and with source‑controlled configuration, CI/CD, release automation, and rollback practices.
  • Strong experience designing and operating observability, logging, monitoring, alerting, and incident‑routing solutions for distributed production platforms.
  • Demonstrated ability to lead major incidents and cross‑functional troubleshooting across infrastructure, applications, data pipelines, and analytics layers.
  • Experience operating production platforms with defined service levels, escalation paths, runbooks, change controls, release processes, and on‑call responsibilities.
  • Strong working knowledge of modern data platforms, including Snowflake, S3-based data lakes, SQL, dbt, managed ingestion tools such as Fivetran or HVR, and custom data pipelines.
  • Understanding of data quality, freshness, lineage, schema evolution, pipeline dependencies, backfills, and recovery procedures.
  • Proficiency in Python, Shell, Bash, or comparable languages for automation and operational tooling.
  • Ability to make well‑reasoned architecture and operational decisions, communicate tradeoffs, estimate work, identify risks, and guide teams through change.
  • Proven experience mentoring engineers, reviewing technical work, delegating ownership, and improving engineering processes across a distributed team.
  • Strong stakeholder‑management and communication skills, including the ability to collect requirements, explain technical risks, and present recommendations to client leaders.
  • Strong written and verbal English communication skills.
  • Availability to work within the LatAm service window of approximately 9:00 AM to 6:00 PM Eastern Time and participate in an agreed senior escalation and on‑call rotation.
Nice to haves
  • Experience with data observability platforms such as SYNQ and operational tooling such as Splunk, PagerDuty, Opsgenie, or comparable solutions.
  • Experience implementing dependency‑aware alerting, selective auto‑remediation, impact analysis, or other advanced reliability practices.
  • Experience leading a managed service, platform operations team, Site Reliability Engineering function, or follow‑the‑sun support model.
  • Experience in healthcare, financial services, or another regulated environment.
What you will do
  • Own the technical direction and end‑to‑end reliability of the managed Data Platform across AWS, Amazon EKS, Kubernetes, Argo Workflows, Snowflake, S3, Fivetran and custom ingestion pipelines, and Tableau dependencies.
  • Lead the technical transition into managed services, including platform discovery, dependency mapping, risk identification, knowledge transfer, shadowing, reverse shadowing, and readiness validation by platform tower.
  • Establish and evolve cloud architecture principles, Infrastructure as Code standards, CI/CD and release practices, observability patterns, operational controls, and platform engineering priorities.
  • Provide technical leadership for AWS infrastructure, Kubernetes and EKS operations, Argo Workflows, deployment automation, secrets and access management, platform capacity, and production reliability.
  • Translate business, service, security, and platform needs into actionable technical requirements, implementation plans, and a prioritized continuous‑improvement backlog.
  • Guide cross‑platform decisions involving Snowflake, dbt, Fivetran, custom ingestion, S3 data‑lake operations, workflow orchestration, data quality, and downstream analytics dependencies.
  • Lead complex cross‑platform troubleshooting and L3 escalation, coordinating engineers when incidents span infrastructure, ingestion, orchestration, Snowflake, and Tableau.
  • Direct technical response during major incidents, including impact assessment, stakeholder communication, recovery strategy, root‑cause analysis, post‑incident review, and preventive actions.
  • Ensure that changes and releases include appropriate technical review, dependency analysis, validation, rollback planning, traceability, and coordination across affected platform components.
  • Define and monitor reliability indicators such as availability, ingestion success, data freshness, workflow reliability, deployment outcomes, alert quality, mean time to restore service, and recurring incident patterns.
  • Lead improvements in automation, observability, auto‑remediation, platform resilience, deployment safety, performance, security, capacity management, and cloud cost efficiency.
  • Partner with Security, Governance, Analytics, IT, and platform stakeholders to ensure least‑privilege access, secrets management, auditability, controlled changes, and appropriate handling of regulated data.
  • Review technical designs and high‑impact changes, challenge assumptions, document tradeoffs, and ensure solutions remain secure, scalable, and supportable within the managed‑service operating model.
  • Mentor Senior and Middle‑level engineers, delegate ownership effectively, improve team practices, and build consistent technical capability across the distributed team.
  • Maintain alignment between the LatAm and India coverage windows through clear ownership, escalation paths, operating procedures, and structured handoffs.
  • Participate in the senior escalation and on‑call model for critical incidents outside staffed service hours.
Perks
  • Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
  • Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
  • Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
  • Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
  • Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
  • Well‑being & support: access local well‑being programs and people‑focused support tailored to your location
Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Data Platform Engineer ID92207
Senior Data Platform Engineer ID92207

AgileEngine, LLC. • Campinas

Híbrido
BRL 731 000 - 993 000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Senior Data Platform Engineer ID92207
Senior Data Platform Engineer ID92207

AgileEngine, LLC. • Curitiba

Teletrabalho
BRL 627 000 - 993 000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Senior Data Platform Engineer ID92207
Senior Data Platform Engineer ID92207

AgileEngine, LLC. • Recife

Teletrabalho
BRL 627 000 - 836 000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Senior Data Platform Engineer ID92207
Senior Data Platform Engineer ID92207

AgileEngine, LLC. • Porto Alegre

Híbrido
BRL 731 000 - 1 097 000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Senior Data Platform Engineer ID92207
Senior Data Platform Engineer ID92207

AgileEngine, LLC. • Rio de Janeiro

Teletrabalho
BRL 731 000 - 993 000
Professional growth
Competitive USD-based compensation
A selection of exciting projects
+1
Staff Engineer (Core & MLOps)
Staff Engineer (Core & MLOps)

Lever, Inc. • Brasil

Teletrabalho
BRL 300 000 - 480 000
Fully remote, remote‑first
Flexible hours
Global collaboration
Senior Snowflake Engineer
Senior Snowflake Engineer

Lever, Inc. • Brasil

Teletrabalho
BRL 250 000 - 450 000
Fully remote
Remote-first
Flexible outcomes
+3
Hiring: Analista De Sistemas (Ti)
Hiring: Analista De Sistemas (Ti)

Agileengine • Belo Horizonte

Teletrabalho
BRL 180 000 - 300 000
Annual learning budget
Remote work with flexible hours
Growth opportunities
+3
DevOps Engineer ID89052
DevOps Engineer ID89052

AgileEngine, LLC. • São Bernardo do Campo

Presencial
BRL 574 000 - 782 000
Professional growth
Competitive USD-based compensation
Exciting projects
+1
DevOps Engineer ID89052
DevOps Engineer ID89052

AgileEngine, LLC. • Florianópolis

Presencial
BRL 120 000 - 180 000
Flextime
Professional growth
Competitive USD-based compensation