DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC.

Metropolitana

Presencial

COP 120.000.000 - 240.000.000

Jornada completa

Hace 8 días

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Growth opportunities
Competitive pay
Remote work
Modern projects
Collaborative culture
Well-being programs

Descripción de la vacante

AgileEngine is seeking a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. You will blend platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz.

You will lead major incidents, own remediation follow-through, and build playbooks that guide response.

This role offers a hands-on, autonomous position with the opportunity to drive systemic improvements across global teams.

Formación

  • 5+ years of experience in DevOps/SRE and incident management.
  • In-depth expertise in multi-cloud defense, IAM, and zero-trust.
  • Hands-on experience with Kubernetes, Terraform, CI/CD, and scripting (Python/Go).
  • Senior-level incident-command experience in a 24x7 production environment.
  • Proven remediation follow-up and post-incident improvements.
  • Ability to draft clear incident communications for technical and executive audiences.
  • Experience creating divisional incident-management playbooks and runbooks.
  • Autonomous and capable of mentoring mid-level SREs.

Responsabilidades

  • Scale and maintain automation across Azure, AWS, and GCP for operational stability.
  • Engineer unified security baselines using IaC (Terraform).
  • Design and optimize enterprise CI/CD pipelines for ASPM ingestion and deployment.
  • Respond to alerts with CSPM tools like Wiz and secure workloads.
  • Act as Incident Commander for major incidents in a 24x7 environment.
  • Own the post-incident loop and drive systemic fixes across groups.
  • Draft and send incident notifications to teams and management.
  • Develop and socialize incident-management playbooks and escalation procedures.

Conocimientos

Kubernetes
Terraform
CI/CD orchestration
Python/Go scripting
incident-command experience
multi-cloud defense
zero-trust principles
autonomy
runbooks automation
incident-notification communications

Herramientas

Wiz
PagerDuty
ServiceNow

Descripción del empleo

DevOps / Site Reliability Engineer ID70127

Full time | AgileEngine | Colombia

Posted On 08/29/2026

Job Information

City Bucaramanga

State/Province Santander

680011

IT Services

Job Description

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

WHY JOIN US

If you're looking for a place to grow, make an impact, and work with people who care, we'd love to meet you!

ABOUT THE ROLE

We are looking for a DevOps / Site Reliability Engineer to maintain operational resilience across Azure, AWS, and GCP in a 24x7 environment. This role blends platform engineering with incident command, using Terraform, CI/CD pipelines, and CSPM tools like Wiz. You will lead major-incident calls, own remediation follow-through, and build the playbooks that guide response.

WHAT YOU WILL DO
  • Scale and maintain the ability to drive operational stability across multi-cloud environments (Azure, AWS, GCP).
  • Engineer unified security policies and configuration baselines using IaC (Terraform) to prevent misconfigurations.
  • Design, maintain, and optimize enterprise CI/CD pipelines to support continuous ASPM ingestion and deployment.
  • Act on continuous monitoring alerts, utilizing Cloud Security Posture Management (CSPM) tools like Wiz to secure workloads.
  • Serve as Incident Commander on major and critical incidents — running the bridge, directing technical workstreams, making time-critical decisions, and coordinating cross-functional responders under pressure.
  • Own the post-incident loop — track remediation items to closure, hold owning teams accountable to timelines, and drive systemic fixes and preventative actions across groups.
  • Draft and send clear, accurate, audience-appropriate incident notifications and status updates to technical teams, management, and stakeholders throughout the incident lifecycle.
  • Develop, maintain, and socialize divisional / group-level incident-management playbooks, runbooks, and escalation procedures that standardize response and reduce time-to-resolution.
MUST HAVES
  • 5+ years of experience.
  • In-depth architectural expertise in multi-cloud defense, federated IAM, and zero-trust principles.
  • Strong practical experience with Kubernetes, Terraform, CI/CD orchestration, and Python/Go scripting.
  • Senior-level, hands-on incident-command experience driving major/critical incident calls to resolution in a 24x7 production environment.
  • Proven track record of remediation follow-up — coordinating with teams and holding owners accountable until issues are fully closed.
  • Demonstrated skill drafting and issuing incident notification communications to both technical and executive audiences.
  • Direct experience authoring divisional/group incident-management playbooks and escalation procedures.
  • Fully autonomous.
  • Drives the architecture of complex automated runbooks and mentors Middle-level SREs.
  • Extensive experience deploying and tuning APIs from modern CNAPP/CSPM platforms, ideally Wiz.
  • Prior experience building platforms subject to strict financial compliance standards (PCI-DSS, SOC2).
NICE TO HAVES
  • PagerDuty — hands‑on experience with on‑call scheduling, alert routing, and incident orchestration.
  • ServiceNow — familiarity with incident, problem, and change management workflows and reporting.
PERKS AND BENEFITS
  • Growth without limits: build your skills through mentorship, internal TechTalks, challenging projects, and a dedicated annual learning budget
  • Competitive compensation: get recognition that reflects your skills and impact, with regular performance and compensation reviews
  • Flexibility: work 100% remotely with flexible hours that support focus, autonomy, and a healthy work rhythm
  • Meaningful, modern projects: build impactful products using modern technologies alongside global teams and leading brands
  • Collaborative culture: join a supportive environment with zero micromanagement where ideas are welcomed and contributions are recognized
  • Well-being & support: access local well-being programs and people-focused support tailored to your location
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Centrosur

Presencial
COP 100.440.000 - 167.400.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
DevOps / Site Reliability Engineer ID70127
DevOps / Site Reliability Engineer ID70127

AgileEngine, LLC. • Sur

Presencial
COP 120.000.000 - 240.000.000
Growth opportunities
Competitive pay
Remote-friendly schedule
+3
Site Reliability Engineer
Site Reliability Engineer

AgileEngine • Colombia

Presencial
COP 283.921.000 - 441.654.000
Professional growth
Competitive USD compensation
Flextime
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Bogotá

Híbrido
COP 285.938.000 - 428.909.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer
DevOps Engineer

Lean Solutions Group • Colombia

Presencial
COP 120.000.000 - 180.000.000
DevOps Engineer ID56470
DevOps Engineer ID56470

AgileEngine • Sur

Híbrido
COP 70.000 - 90.000
Mentorship
TechTalks
USD-based compensation
+2
Remote Senior DevOps & SRE — Multi-Cloud Incident Lead
Remote Senior DevOps & SRE — Multi-Cloud Incident Lead

AgileEngine, LLC. • Perímetro Urbano Barranquilla

Presencial
COP 167.400.000 - 290.160.000
Growth opportunities
Competitive salary
Remote-friendly
+3
DevOps Engineer ID56470
DevOps Engineer ID56470

AgileEngine • Perímetro Urbano Barranquilla

Híbrido
COP 292.440.000 - 438.661.000
Professional growth
Competitive compensation
Exciting projects
+1
DevOps Engineer ID38563 ($3,000 signing bonus)
DevOps Engineer ID38563 ($3,000 signing bonus)

AgileEngine • Perímetro Urbano Manizales

Híbrido
COP 241.409.000 - 321.880.000
Professional growth
Competitive compensation
Exciting projects
+1