DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC.

Ciudad de México

Presencial

MXN 900.000 - 1.200.000

Jornada completa

Hace 7 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Growth without limits
Competitive compensation
Flexibility: 100% remote
Meaningful, modern projects
Collaborative culture
Well‑being & support

Descripción de la vacante

AgileEngine is seeking a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. You will lead major incident calls, own remediation, and build playbooks guiding response.

You will architect automated runbooks, socializing playbooks within the division and ensuring timely communications to stakeholders. Remote-friendly, with a focus on security and reliability.

Formación

  • 5+ years of hands-on DevOps/SRE experience in multi-cloud environments.
  • Strong architectural expertise in multi-cloud defense and zero-trust.
  • Proficient with Kubernetes, Terraform, CI/CD orchestration, and scripting in Python/Go.
  • Senior incident-command experience in 24x7 production settings.
  • Experience drafting incident communications and runbooks.

Responsabilidades

  • Scale and maintain operational stability across Azure, AWS, and GCP.
  • Engineer IaC policies and baselines to prevent misconfigurations.
  • Design and optimize enterprise CI/CD pipelines for ASPM ingestion and deployment.
  • Lead major incidents as Incident Commander and coordinate responders.
  • Own post-incident remediation tracking and cross-team accountability.
  • Draft clear incident notifications for technical and executive audiences.

Conocimientos

Kubernetes
Terraform
CI/CD orchestration
Python/Go scripting
Incident-command experience
Multi-cloud defense
Zero-trust principles
Wiz CSPM

Herramientas

Wiz CSPM
CI/CD tooling
Azure/AWS/GCP environments

Descripción del empleo

DevOps / Site Reliability Engineer ID70127

Full time | AgileEngine | Mexico

Posted On 08/29/2026

Job Information

City Ciudad de México

State/Province México

01210

IT Services

Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. This role blends platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz. You will lead major-incident calls, own remediation follow-through, and build the playbooks that guide response.

WHAT YOU WILL DO
  • - Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP).
  • - Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations.
  • - Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment.
  • - Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads.
  • - Serve as Incident Commander on major and critical incidents — running the bridge, directing technical workstreams, making time‑critical decisions, and coordinating cross‑functional responders under pressure.
  • - Own the post‑incident loop — track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups.
  • - Draft and send clear, accurate, audience‑appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle.
  • - Develop, maintain, and socialize divisional / group‑level incident‑management playbooks, runbooks, and escalation procedures that standardize response and reduce time‑to‑resolution.
MUST HAVES
  • - 5+ years of experience.
  • - In-depth architectural expertise in multi‑cloud defense, federated IAM, and zero‑trust principles.
  • - Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting.
  • - Senior‑level, hands‑on incident‑command experience driving major/critical incident calls to resolution in a 24x7 production environment.
  • - Proven track record of remediation follow‑up — coordinating with teams and holding owners accountable until issues are fully closed.
  • - Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences.
  • - Direct experience authoring divisional/group incident‑management playbooks and escalation procedures.
  • - Fully autonomous.
  • - Drives the architecture of complex automated runbooks and mentors Middle‑level SREs.
  • - Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz.
  • - Prior experience building platforms subject to strict financial compliance standards ( PCI‑DSS, SOC2 ).
NICE TO HAVES
  • - PagerDuty — hands‑on experience with on‑call scheduling, alert routing, and incident orchestration.
  • - ServiceNow — familiarity with incident, problem, and change management workflows and reporting.
PERKS AND BENEFITS
  • - Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
  • - Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
  • - Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
  • - Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
  • - Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
  • - Well‑being & support: access local well‑being programs and people‑focused support tailored to your location
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Monterrey

Presencial
MXN 607.000 - 1.036.000
Growth opportunities
Competitive compensation
Remote work with flexible hours
+3
DevOps Engineer (Senior/Lead) ID55632
DevOps Engineer (Senior/Lead) ID55632

AgileEngine • Ciudad de México

Híbrido
MXN 1.252.000 - 1.611.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer (Senior/Lead) ID55632
DevOps Engineer (Senior/Lead) ID55632

AgileEngine • Santiago de Querétaro

Híbrido
MXN 400.000 - 700.000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Rosarito

Híbrido
MXN 870.000 - 1.306.000
Professional growth
Competitive compensation
Exciting projects
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Monterrey

Híbrido
MXN 1.568.000 - 2.091.000
Professional growth
Competitive compensation
Exciting projects
+1
Remote DevOps & SRE - Multi-Cloud Reliability Lead
Remote DevOps & SRE - Multi-Cloud Reliability Lead

AgileEngine, LLC. • Monterrey

Presencial
MXN 607.000 - 1.036.000
Growth opportunities
Competitive compensation
Remote work with flexible hours
+3
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Santiago de Querétaro

Híbrido
MXN 1.394.000 - 1.918.000
Professional growth programs
Competitive compensation
Exciting projects with top-tier clients
+1
Senior Multi-Cloud SRE & Incident Commander (Remote)
Senior Multi-Cloud SRE & Incident Commander (Remote)

AgileEngine, LLC. • Santiago de Querétaro

Presencial
MXN 772.000 - 924.000
Growth opportunities
Competitive compensation
100% remote with flexible hours
+3
Site Reliability Engineer ID60188
Site Reliability Engineer ID60188

AgileEngine • Ciudad de México

Híbrido
MXN 1.049.000 - 1.400.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior Multi-Cloud SRE & Incident Commander — Remote
Senior Multi-Cloud SRE & Incident Commander — Remote

AgileEngine, LLC. • Región Centro

Presencial
MXN 1.000.000 - 1.600.000
Growth opportunities
Remote work with flexible hours
Annual learning budget