Senior Site Reliability Engineer

AgileEngine

Colombia

Híbrido

COP 275.668.000 - 367.557.000

Jornada completa

Hace 6 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Professional growth
Competitive compensation
A selection of exciting projects
Flextime

Descripción de la vacante

AgileEngine in Colombia is seeking a Senior Site Reliability Engineer to own core system administration for on-premise and SaaS-hosted environments, with a strong focus on Kubernetes, monitoring, and observability using Snowflake and OpenTelemetry.

You will participate in on-call rotations, incident response, RCA, and automate infrastructure tasks with Bash, Python, or Go, ensuring health across ESM and ECP platforms.

Formación

  • 4+ years of infrastructure management experience.
  • Experience with Kubernetes.
  • Experience with monitoring and observability.
  • Familiarity with Snowflake and OpenTelemetry ecosystems.
  • Strong background in Linux/Unix administration.
  • Proficiency in scripting languages (Bash, Python, or Go).

Responsabilidades

  • Provide core system administration for enterprise infrastructure.
  • Maintain, scale, and ensure operational stability of on-premise and SaaS-hosted systems.
  • Manage and scale containerized environments using Kubernetes.
  • Participate in on-call rotations, incident response, and root-cause analysis (RCA).
  • Automate repetitive infrastructure tasks using scripting and IaC.

Conocimientos

Kubernetes
Linux/Unix
SRE tasks
Snowflake
OpenTelemetry
Bash
Python
Go
On-call rotations
Incident response
IaC

Herramientas

Terraform
Ansible

Descripción del empleo

We are looking for a Senior Site Reliability Engineer to provide core system administration and operational stability for enterprise on-premise and SaaS-hosted systems, with a strong focus on Kubernetes cluster management, monitoring, and observability using Snowflake and OpenTelemetry. You will participate in on-call rotations, incident response, and root-cause analysis, automate infrastructure tasks using Python, Bash, or Go, and ensure system health across ESM and ECP platform environments. Experience with service mesh architectures is highly valued for ECP-focused roles.

What you will do

  • Provide core system administration for the enterprise infrastructure.
  • Focus on the maintenance, scaling, and operational stability of various on-premise and SaaS-hosted systems.
  • Manage and scale containerized environments using Kubernetes.
  • For ECP-focused roles: Lean heavily into executing monitoring and observability tasks to ensure system health.
  • Participate in on-call rotations, incident response, and root-cause analysis (RCA).
  • Automate repetitive infrastructure tasks using scripting and infrastructure-as-code (IaC).

Must haves

  • 4+ years of infrastructure management experience.
  • Experience working with Kubernetes.
  • Experience with monitoring, observability, and related SRE tasks.
  • Familiarity with Snowflake and OpenTelemetry ecosystems.
  • Strong background in Linux/Unix administration.
  • Proficiency in scripting languages (e.g., Bash, Python, or Go).

Nice to haves

  • For ECP SREs: Experience with Service Mesh architectures is highly ideal.
  • Experience with Infrastructure as Code tools such as Terraform or Ansible.
  • Familiarity with cloud platforms (AWS, GCP, or Azure).

Perks and Benefits

  • Professional growth

Accelerate your professional journey with mentorship, TechTalks, and personalized growth roadmaps

  • Competitive compensation

We match your ever-growing skills, talent, and contributions with competitive USD-based compensation and budgets for education, fitness, and team activities

  • A selection of exciting projects

Join projects with modern solutions development and top-tier clients that include Fortune 500 enterprises and leading product brands

  • Flextime

Tailor your schedule for an optimal work-life balance, by having the options of working from home and going to the office – whatever makes you the happiest and most productive.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Remote Senior SRE - Kubernetes, Observability & Automation
Remote Senior SRE - Kubernetes, Observability & Automation

AgileEngine • Capital

Presencial
COP 120.000.000 - 240.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MPS Group LLC • Bogotá

Presencial
COP 200.880.000 - 312.480.000
Senior SRE: Kubernetes, Observability & Flexible Hours
Senior SRE: Kubernetes, Observability & Flexible Hours

AgileEngine • Colombia

Híbrido
COP 275.668.000 - 367.557.000
Professional growth
Competitive compensation
A selection of exciting projects
+1
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior SRE - Kubernetes & Observability (Remote)
Senior SRE - Kubernetes & Observability (Remote)

INGEPSY • Risaralda

Presencial
COP 436.042.000 - 591.771.000
Growth without limits
Competitive compensation
100% remote with flexible hours
+3
Remote Senior SRE - Kubernetes & Observability
Remote Senior SRE - Kubernetes & Observability

INGEPSY • Sucre

Presencial
COP 373.750.000 - 498.334.000
Growth without limits
Competitive compensation
Flexibility — remote work with flexi‑s
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Publicis Sapient • Colombia

Presencial
COP 284.270.000 - 473.784.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LanceSoft, Inc. • Colombia

Presencial
COP 90.000.000 - 150.000.000
Senior Site Reliability Engineer - Remote & Impactful
Senior Site Reliability Engineer - Remote & Impactful

AgileEngine • Bogotá

Presencial
COP 200.880.000 - 334.800.000
Growth opportunities
Competitive compensation
Remote work 100%
+3
Senior SRE: Remote Kubernetes, Observability & Automation
Senior SRE: Remote Kubernetes, Observability & Automation

AgileEngine • Bogotá

Presencial
COP 120.000.000 - 180.000.000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3