DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC.

Pereira

Presencial

COP 133.920.000 - 178.560.000

Jornada completa

Hace 11 días

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Growth without limits
Competitive compensation
Remote work
Meaningful, modern projects
Collaborative culture
Well-being programs

Descripción de la vacante

AgileEngine is seeking a DevOps / Site Reliability Engineer to maintain 24/7 operational resilience for a multi-cloud enterprise security program. You will act as Incident Commander during major incidents, own IaC, CI/CD pipelines, and CSPM telemetry using Terraform and Wiz, driving remediation and stakeholder communications.

The role requires 5+ years in SRE with hands-on incident command in financial services environments, and experience with PCI-DSS/SOC2 compliance is a plus.

Formación

  • 5+ years of experience in SRE or DevOps in 24/7 production.
  • Deep multi-cloud defense, federated IAM, and zero-trust principles.
  • Strong Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting experience.
  • Senior-level, hands-on incident-command experience in a 24x7 environment.
  • Proven remediation follow-up and ownership across teams.
  • Experience drafting incident communications for technical and executive audiences.
  • Experience authoring incident-management playbooks and escalation procedures.
  • Autonomous with ability to mentor mid-level SREs.
  • Experience deploying APIs via CNAPP/CSPM platforms (Wiz).
  • Experience with PCI-DSS and SOC2 compliance.

Responsabilidades

  • Scale and maintain stability across multi-cloud environments (Azure, AWS, GCP).
  • Engineer IaC-based security baselines to prevent misconfigurations.
  • Design and optimize CI/CD pipelines to support ingestion and deployment.
  • Respond to alerts using CSPM tools to secure workloads.
  • Act as Incident Commander for major incidents and coordinate responders.
  • Lead post-incident remediation tracking and timelines to closure.
  • Draft clear incident notifications for technical teams and executives.
  • Develop and socialize incident-management playbooks and runbooks.

Conocimientos

Kubernetes
Terraform
CI/CD orchestration
Python/Go scripting
Incident command
Multi-cloud architecture
IAM / zero-trust

Herramientas

Wiz
CNAPP/CSPM platforms

Descripción del empleo

DevOps / Site Reliability Engineer ID70127

Full time | AgileEngine | Colombia

Posted On 08/11/2026

Job Information

City Pereira

State/Province Risaralda

660004

IT Services

Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience and 24/7 stability for a multi-cloud enterprise security program, serving as Incident Commander on major and critical incidents while also owning IaC, CI/CD pipelines, and CSPM telemetry. You will drive major-incident calls, own post-incident remediation follow-through, draft stakeholder communications, and develop divisional incident-management playbooks alongside multi-cloud security guardrails using Terraform and Wiz. The role requires 5+ years of SRE experience with hands-on incident command in a 24/7 financial services environment.

WHAT YOU WILL DO
  • Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP).
  • Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations.
  • Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment.
  • Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads.
  • Serve as Incident Commander on major and critical incidents — running the bridge, directing technical workstreams, making time‑critical decisions, and coordinating cross‑functional responders under pressure.
  • Own the post‑incident loop — track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups.
  • Draft and send clear, accurate, audience‑appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle.
  • Develop, maintain, and socialize divisional / group‑level incident‑management playbooks, runbooks, and escalation procedures that standardize response and reduce time‑to‑resolution.
MUST HAVES
  • 5+ years of experience.
  • In-depth architectural expertise in multi‑cloud defense, federated IAM, and zero‑trust principles.
  • Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting.
  • Senior‑level, hands‑on incident‑command experience driving major/critical incident calls to resolution in a 24x7 production environment.
  • Proven track record of remediation follow‑up — coordinating with teams and holding owners accountable until issues are fully closed.
  • Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences.
  • Direct experience authoring divisional/group incident‑management playbooks and escalation procedures.
  • Fully autonomous.
  • Drives the architecture of complex automated runbooks and mentors Middle‑level SREs.
  • Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz.
  • Prior experience building platforms subject to strict financial compliance standards (PCI-DSS, SOC2).
NICE TO HAVES
  • PagerDuty — hands‑on experience with on‑call scheduling, alert routing, and incident orchestration.
  • ServiceNow — familiarity with incident, problem, and change management workflows and reporting.
PERKS AND BENEFITS
  • Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
  • Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
  • Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
  • Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
  • Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
  • Well‑being & support: access local well‑being programs and people‑focused support tailored to your location
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Metropolitana

Presencial
COP 90.000.000 - 150.000.000
Growth opportunities
Competitive pay
Remote work with flexible hours
+2
Senior DevOps & SRE — Remote, Multi-Cloud & Incident Command
Senior DevOps & SRE — Remote, Multi-Cloud & Incident Command

AgileEngine, LLC. • Metropolitana

Presencial
COP 90.000.000 - 150.000.000
Growth opportunities
Competitive pay
Remote work with flexible hours
+2
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Remote Senior DevOps & SRE: Multi-Cloud Reliability Lead
Remote Senior DevOps & SRE: Multi-Cloud Reliability Lead

AgileEngine, LLC. • Pereira

Presencial
COP 133.920.000 - 178.560.000
Growth without limits
Competitive compensation
Remote work
+3
Remote DevOps & SRE Lead — Multi-Cloud & Incident Command
Remote DevOps & SRE Lead — Multi-Cloud & Incident Command

AgileEngine, LLC. • Centrosur

Presencial
COP 133.920.000 - 223.200.000
Growth opportunities
Competitive compensation
Remote work 100% remote
+3
Senior DevOps & SRE - Incident Command, Multi-Cloud
Senior DevOps & SRE - Incident Command, Multi-Cloud

AgileEngine, LLC. • Medellín

Presencial
COP 120.000.000 - 160.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote
+3
Senior DevOps & SRE — Multi-Cloud Incident Commander
Senior DevOps & SRE — Multi-Cloud Incident Commander

AgileEngine, LLC. • Cartagena de Indias

Presencial
COP 60.000.000 - 90.000.000
Growth opportunities
Competitive pay
Remote work
+3
Senior Infrastructure Engineer ID82552
Senior Infrastructure Engineer ID82552

AgileEngine • Centrosur

A distancia
COP 60.000.000 - 90.000.000
Growth opportunities
Competitive compensation
Remote work
+3
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Bogotá

Híbrido
COP 285.938.000 - 428.909.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer ID38563 ($3,000 signing bonus)
DevOps Engineer ID38563 ($3,000 signing bonus)

AgileEngine • Perímetro Urbano Manizales

Híbrido
COP 241.409.000 - 321.880.000
Professional growth
Competitive compensation
Exciting projects
+1