SRE Software Engineer: Build Resilient Cloud & Kubernetes

Jobtailor

Bogotá

Presencial

COP 100.000.000 - 180.000.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Jobtailor in Bogotá, Colombia is seeking a Site Reliability Engineer to join our operations team. You will support and maintain production environments, manage Kubernetes clusters, and optimize cloud infrastructure across AWS/Azure/GCP.

The role emphasizes reliability, incident response, RCA, and continuous improvement in collaboration with engineering and product teams. You will implement IaC, automate deployments, develop observability dashboards in Prometheus/Grafana, and contribute to

Formación

  • 4+ years of experience in Site Reliability Engineering, DevOps, Cloud Operations, or Infrastructure Engineering.
  • Strong hands-on experience with Linux administration, troubleshooting, and production support.
  • Experience managing and supporting Kubernetes and containerized workloads (Docker/OpenShift is a plus).
  • Solid knowledge of AWS, Azure, or GCP cloud environments.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or CloudWatch.
  • Experience with Infrastructure as Code (Terraform preferred) and CI/CD pipelines.
  • Ability to troubleshoot complex production issues, perform Root Cause Analysis (RCA), and drive preventive improvements.
  • Working knowledge of automation and scripting using Bash, Python, or Go.
  • Intermediate to advanced English (B2+).

Responsabilidades

  • Support and maintain business-critical production environments, ensuring high availability and system reliability.
  • Monitor infrastructure and applications, proactively identifying and resolving issues before they impact users.
  • Participate in incident response activities, troubleshooting production outages and coordinating recovery efforts.
  • Perform RCA and contribute to postmortems, corrective actions, and continuous improvement initiatives.
  • Manage and optimize Kubernetes clusters and cloud infrastructure.
  • Develop and maintain monitoring dashboards, alerts, and observability solutions.
  • Automate operational processes and infrastructure deployments using IaC and scripting.
  • Collaborate with engineering and product teams to improve scalability, performance, and operational excellence.
  • Support and enhance CI/CD pipelines to ensure reliable and efficient software delivery.

Conocimientos

Site Reliability Engineering
Kubernetes Management
Cloud Infrastructure (AWS, Azure, GCP)
Infrastructure as Code (Terraform)
Monitoring & Observability (Prometheus
Grafana & Datadog
CI/CD Pipelines
Automation and Scripting (Bash, Python

Herramientas

Kubernetes
Docker
OpenShift
Prometheus
Grafana
Datadog
Splunk
ELK
CloudWatch

Descripción del empleo

Jobtailor in Bogotá, Colombia is seeking a Site Reliability Engineer to join our operations team. You will support and maintain production environments, manage Kubernetes clusters, and optimize cloud infrastructure across AWS/Azure/GCP.

The role emphasizes reliability, incident response, RCA, and continuous improvement in collaboration with engineering and product teams. You will implement IaC, automate deployments, develop observability dashboards in Prometheus/Grafana, and contribute to

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

SRE: Build Resilient Cloud Systems
SRE: Build Resilient Cloud Systems

Capgemini Engineering • Bogotá

Presencial
COP 90.000.000 - 150.000.000
SRE Software Engineer
SRE Software Engineer

Capgemini • Bogotá

Híbrido
Senior SRE - Remote, Kubernetes & Reliability
Senior SRE - Remote, Kubernetes & Reliability

BairesDev • Colombia

Presencial
COP 286.250.000 - 413.473.000
100% remote work
Flexible hours
Competitive compensation
Site Reliability Engineer - Cloud Reliability (Remote/Flex)
Site Reliability Engineer - Cloud Reliability (Remote/Flex)

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Service Reliability Engineer
Service Reliability Engineer

1083 Amadeus IT Group Colombia, S.A.S. • Colombia

Presencial
COP 182.089.000 - 254.926.000
Competitive remuneration
Vacation and holiday paid time off
Health insurances
+3
Remote SRE Engineer: Cloud Reliability & Kubernetes
Remote SRE Engineer: Cloud Reliability & Kubernetes

Encora Inc. • Bogotá

Presencial
COP 281.171.000 - 406.136.000
Site Reliability Engineer: Build Resilient Cloud & Automation
Site Reliability Engineer: Build Resilient Cloud & Automation

CINTE Colombia • Colombia

Presencial
COP 45.000 - 60.000
Senior Site Reliability Engineer - Cloud & Observability
Senior Site Reliability Engineer - Cloud & Observability

cobre.co • Colombia

A distancia
Remote SRE - 24/7 On-Call, DevOps & Reliability
Remote SRE - 24/7 On-Call, DevOps & Reliability

DCT • Bogotá

A distancia
COP 156.225.000 - 234.339.000
Senior SRE: Live Production Reliability & Automation
Senior SRE: Live Production Reliability & Automation

LanceSoft, Inc. • Colombia

Presencial
COP 90.000.000 - 150.000.000