DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC.

Sur

Presencial

COP 120.000.000 - 240.000.000

Jornada completa

Hace 8 días

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Growth opportunities
Competitive pay
Remote-friendly schedule
Meaningful projects
Collaborative culture
Well-being programs

Descripción de la vacante

AgileEngine, a leading IT services firm in Colombia, seeks aDevOps / Site Reliability Engineer to safeguard 24x7 operations across Azure, AWS, and GCP. You will lead major incidents, craft runbooks, and drive remediation while shaping security baselines with IaC and CSPM tooling.

The role emphasizes autonomous leadership, complex automation, and collaboration with global teams on modern cloud-native platforms. Cali-based, on-site with potential for remote flex in some cases.

Formación

  • 5+ years of experience in DevOps/SRE, multi-cloud environments.
  • In-depth knowledge of multi-cloud defense, federated IAM, and zero-trust principles.
  • Hands-on with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting.
  • Senior-level incident-command experience in 24x7 production environments.
  • Proven remediation follow-up and ownership until closure.
  • Experience drafting incident notifications for technical and executive audiences.
  • Experience authoring divisional/group incident-management playbooks and escalation procedures.
  • Fully autonomous with ability to mentor mid-level SREs.
  • Experience deploying and tuning APIs from CNAPP/CSPM platforms, Wiz preferred.
  • Experience with PCI-DSS and SOC2 compliance.

Responsabilidades

  • Scale and maintain operational stability across Azure, AWS, and GCP.
  • Engineer security policies and configuration baselines using Terraform.
  • Design and optimize enterprise CI/CD pipelines for ASPM ingestion and deployment.
  • Respond to monitoring alerts using CSPM tools to secure workloads.
  • Act as Incident Commander during major incidents and coordinate cross-functional responders.
  • Own post-incident remediation loops and timelines for closure.
  • Draft and send incident notifications to technical teams, management, and stakeholders.
  • Develop and socialize incident-management playbooks and runbooks across groups.

Conocimientos

Multi-cloud defense
IAM and zero-trust
Kubernetes
Terraform
CI/CD orchestration
Python/Go scripting
Incident command leadership
Remediation follow-up
Playbooks & runbooks
CNAPP/CSPM (Wiz)
PCI-DSS / SOC2 experience

Herramientas

Wiz CSPM

Descripción del empleo

DevOps / Site Reliability Engineer ID70127

Full time | AgileEngine | Colombia

Posted On 08/29/2026

Job Information

City Cali

State/Province Valle del Cauca

760004

IT Services

Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. This role blends platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz. You will lead major-incident calls, own remediation follow-through, and build the playbooks that guide response.

WHAT YOU WILL DO
  • - Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP).
  • - Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations.
  • - Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment.
  • - Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads.
  • - Serve as Incident Commander on major and critical incidents — running the bridge, directing technical workstreams, making time-critical decisions, and coordinating cross-functional responders under pressure.
  • - Own the post-incident loop — track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups.
  • - Draft and send clear, accurate, audience-appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle.
  • - Develop, maintain, and socialize divisional / group-level incident-management playbooks, runbooks, and escalation procedures that standardize response and reduce time-to-resolution.
MUST HAVES
  • - 5+ years of experience.
  • - In-depth architectural expertise in multi-cloud defense, federated IAM, and zero-trust principles.
  • - Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting.
  • - Senior-level, hands-on incident-command experience driving major/critical incident calls to resolution in a 24x7 production environment.
  • - Proven track record of remediation follow-up — coordinating with teams and holding owners accountable until issues are fully closed.
  • - Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences.
  • - Direct experience authoring divisional/group incident-management playbooks and escalation procedures.
  • - Fully autonomous.
  • - Drives the architecture of complex automated runbooks and mentors Middle-level SREs.
  • - Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz.
  • - Prior experience building platforms subject to strict financial compliance standards (PCI-DSS, SOC2).
NICE TO HAVES
  • - PagerDuty — hands-on experience with on-call scheduling, alert routing, and incident orchestration.
  • - ServiceNow — familiarity with incident, problem, and change management workflows and reporting.
PERKS AND BENEFITS
  • - Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
  • - Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
  • - Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
  • - Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
  • - Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
  • - Well-being & support: access local well-being programs and people-focused support tailored to your location
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Metropolitana

Presencial
COP 120.000.000 - 240.000.000
Growth opportunities
Competitive pay
Remote work
+3
DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Centrosur

Presencial
COP 100.440.000 - 167.400.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Site Reliability Engineer
Site Reliability Engineer

AgileEngine • Colombia

Presencial
COP 283.921.000 - 441.654.000
Professional growth
Competitive USD compensation
Flextime
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Bogotá

Híbrido
COP 285.938.000 - 428.909.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer ID56470
DevOps Engineer ID56470

AgileEngine • Perímetro Urbano Barranquilla

Híbrido
COP 292.440.000 - 438.661.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer ID56470
DevOps Engineer ID56470

AgileEngine • Sur

Híbrido
COP 70.000 - 90.000
Mentorship
TechTalks
USD-based compensation
+2
Remote Senior DevOps & SRE — Multi-Cloud Incident Lead
Remote Senior DevOps & SRE — Multi-Cloud Incident Lead

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Presencial
COP 167.400.000 - 290.160.000
Growth opportunities
Competitive salary
Remote-friendly
+3
Remote Senior DevOps & SRE — Multi-Cloud & Incident Command
Remote Senior DevOps & SRE — Multi-Cloud & Incident Command

AgileEngine, LLC. • Centrosur

Presencial
COP 100.440.000 - 167.400.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
DevOps Engineer ID38563 ($3,000 signing bonus)
DevOps Engineer ID38563 ($3,000 signing bonus)

AgileEngine • Perímetro Urbano Manizales

Híbrido
COP 241.409.000 - 321.880.000
Professional growth
Competitive compensation
Exciting projects
+1