Service Reliability Engineer

1083 Amadeus IT Group Colombia, S.A.S.

Colombia

Presencial

COP 182.089.661 - 254.925.525

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Competitive remuneration
Vacation and holiday paid time off
Health insurances
Hybrid work model
Professional development opportunities
Diverse and inclusive workplace

Descripción de la vacante

1083 Amadeus IT Group Colombia, S.A.S. is seeking a Service Reliability Engineer to ensure the reliability and scalability of mission-critical airline platforms. The role involves automation, performance tracking, and teamwork across globally distributed teams.

The ideal candidate will have experience in production operations, strong technical skills, and good English communication skills. This position offers a hybrid work model in Bogotá with competitive benefits.

Join a diverse workplace that promotes professional development and values inclusion.

Formación

  • 3+ years of experience as a Site Reliability Engineer or similar role.
  • Strong experience in production operations and incident management.
  • Good communication skills in English.

Responsabilidades

  • Ensure high availability and resilience of production systems.
  • Automate operational tasks and improve system reliability.
  • Build and enhance monitoring frameworks.

Conocimientos

Site Reliability Engineering
Incident Management
Production Operations
Linux
Cloud Platforms
Python
Kubernetes
Observability Tools
Communication Skills

Herramientas

Grafana
Prometheus
ELK
Splunk
Ansible

Descripción del empleo

Service Reliability Engineer

Location: Bogotá

Open to SRE's with different technical backgrounds and all levels of experience, including cloud-native, platform, and application-focused reliability engineering.

Summary of the role

We are looking for Site Reliability Engineers (SREs) to join our global teams supporting mission‑critical airline platforms and systems. In this role, you will focus on ensuring system reliability, availability, scalability, and performance, while driving automation and operational excellence across distributed environments. You will collaborate with globally distributed teams in a follow‑the‑sun model, supporting production systems and continuously improving reliability and operational processes.

Key Responsibilities
  • Reliability & Production Operations: Ensure high availability, scalability, and resilience of production systems; Conduct incident, problem, and change management following ITIL practices; Perform root cause analysis (RCA) and drive resolution of production issues; Support systems across multiple environments (test, staging, production); Participate in on‑call / follow‑the‑sun rotations.
  • Automation & Reliability Engineering: Automate repetitive operational tasks, deployments, and recovery processes; Improve system reliability through engineering solutions; Contribute to continuous improvement of operational processes and efficiency; Support implementation of deployment strategies (e.g., blue/green, canary).
  • Observability & Performance: Build and enhance monitoring, alerting, and observability frameworks; Improve visibility across metrics, logs, and traces; Track and improve SLOs/SLAs and system performance; Perform proactive reliability analysis and capacity planning.
  • Infrastructure & Platform Reliability: Operate and support systems across cloud and on‑prem environments; Work with containerized and distributed systems; Support Linux and/or Windows‑based environments; Contribute to system architecture improvements, resilience, and scalability.
  • Application Reliability (Multi‑stack): Troubleshoot distributed systems, microservices, and APIs; Diagnose issues using logs, monitoring tools, and profiling techniques; Support applications across different stacks, such as .NET/C# applications and other backend technologies; Manage runtime configuration, dependencies, and system health.
  • Collaboration & Continuous Improvement: Partner closely with engineering, platform, and product teams; Contribute to post‑incident reviews and reliability improvements; Document processes, playbooks, and troubleshooting guides; Support knowledge sharing and mentoring of junior engineers.
Required Skills & Experience
  • Core SRE Capabilities: Experience as a Site Reliability Engineer or in a similar production engineering role (typically 3+ years, adaptable based on seniority); Strong experience in production operations and incident management; Solid understanding of reliability concepts (availability, latency, scalability, resilience); Experience working in mission‑critical environments.
  • Technical Skills: Operating systems: Linux and/or Windows Server; Observability tools: Grafana, Prometheus, ELK, Splunk or similar; CI/CD and automation tooling (Jenkins, GitHub Actions, etc.); Cloud platforms: Azure, AWS, or GCP; Scripting/programming: Python, Shell, Go, or similar; Understanding of distributed systems, microservices, and databases; Familiarity with ITIL processes; Container orchestration: Kubernetes, OpenShift; .NET/C# application debugging and IIS administration; Automation tools: Ansible or similar; Event streaming (Kafka), caching, or messaging systems; Database technologies (SQL/NoSQL).
  • Soft Skills: Good communication skills in English; Strong ownership and reliability mindset; Ability to perform under pressure in production environments; Strong collaboration and communication skills across global teams; Continuous improvement and problem‑solving mindset.
Benefits
  • Competitive remuneration and individual and company annual bonus.
  • Vacation and holiday paid time off.
  • Health insurances and other competitive benefits.
  • Hybrid work at our Bogotá office.
  • Professional development with online learning hubs covering technical and soft skills.
  • Diverse and inclusive workplace.
Equal Opportunity

Amadeus is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to gender, race, ethnicity, sexual orientation, age, beliefs, disability or any other characteristics protected by law.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Global SRE Engineer - Reliability, Automation & Observability
Global SRE Engineer - Reliability, Automation & Observability

1083 Amadeus IT Group Colombia, S.A.S. • Bogotá

Híbrido
COP 182.089.000 - 254.926.000
Competitive remuneration
Vacation and holiday paid time off
Health insurances
+3
Senior SRE - Hybrid, Global DevOps & Reliability
Senior SRE - Hybrid, Global DevOps & Reliability

Amadeus • Bogotá

Híbrido
COP 90.000.000 - 150.000.000
Hybrid work in Bogota
Annual bonus
Health insurance
+1
Service Reliability Engineer
Service Reliability Engineer

Amadeus • Colombia

Presencial
COP 258.923.000 - 332.902.000
Competitive remuneration
Annual bonuses
Health insurance
+2
Senior Service Reliability Engineer
Senior Service Reliability Engineer

Amadeus Hospitality • Bogotá

Híbrido
COP 60.000.000 - 90.000.000
Annual bonus
Health insurance
Paid time off
+1
Senior Service Reliability Engineer
Senior Service Reliability Engineer

1083 Amadeus IT Group Colombia, S.A.S. • Bogotá

Híbrido
COP 180.000.000 - 320.000.000
Annual bonus
Health insurance
Site Reliability Engineer
Site Reliability Engineer

DCT • Bogotá

A distancia
COP 156.225.000 - 234.339.000
Career Growth & Mentorship
Flexible Work Environment
Generative & Collaborative Culture
Ingeniero DevOps / Site Reliability Engineer (SRE)
Ingeniero DevOps / Site Reliability Engineer (SRE)

Gopass • Bogotá

Híbrido
COP 72.000.000 - 120.000.000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Michael Page Colombia • San Gil

Presencial
COP 228.815.000 - 343.224.000
Crecimiento profesional a través de desafíos técnicos
Trabajo con tecnologías cloud de vanguardia
Exposición a prácticas SRE modernas
SRE Software Engineer
SRE Software Engineer

Capgemini • Bogotá

Híbrido
Service Reliability Engineer
Service Reliability Engineer

Amadeus Hospitality • Bogotá

Híbrido
COP 90.000.000 - 150.000.000
Hybrid work model
Health insurance
Annual bonus
+1