Service Reliability Engineer

Amadeus

Bogotá

Híbrido

COP 70.000.000 - 110.000.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Hybrid working model
Annual bonus
Online training hubs
Diverse and inclusive workplace

Descripción de la vacante

Amadeus in Bogotá is seeking a Service Reliability Engineer to guard critical data pipelines and batch workloads on a 24x7 frontline team. You will monitor high-volume processing jobs, triage alerts, and execute interventions to meet business deadlines, while automating repetitive tasks and improving pipeline resilience.

The role emphasizes incident handling, post-incident analysis, and close collaboration with application owners to prevent recurrence and ensure smooth operations across cloud

Formación

  • Bachelor’s degree in CS or related field or equivalent practical experience.
  • Experience in production support or operations, ideally on a 24x7 rotation.
  • Familiarity with batch processing and job scheduling (e.g., Control-M, SFTP).
  • Familiarity with batch scripting and system architecture.
  • Ability to work independently and under pressure.

Responsabilidades

  • Run and monitor scheduled batch jobs.
  • Troubleshoot system errors and performance issues.
  • Do incident management to recover incidents and collaborate with application owners on problem management to understand the root cause and avoid reoccurrence.
  • Maintain documentation and logs and work to improve processes (feedback loops).
  • Track metrics and KPIs on batch support activities.
  • Coordinate with other departments to ensure job dependencies are met.
  • Automate routine tasks and deployments using scripting and configuration management tools (e.g., Bash, Python, Ansible, Terraform).

Conocimientos

Linux
Windows
Cloud
Batch processing
Python
Monitoring
Incident management
Automation
SRE practices

Educación

Bachelor’s degree in CS or related field

Herramientas

Prometheus
Grafana
Splunk
Bash
Python
Ansible
Terraform

Descripción del empleo

About The Business Area/Department

The

Job Title

Service Reliability Engineer

About The Business Area/Department

The Platform & Network Support Services department provides reliable, continuous, and globally coordinated frontline support for network, batches, cloud platforms, and critical infrastructure services as well as monitoring services, safeguarding operational stability and enabling seamless system lifecycle operations.

Summary Of The Role

A Service Reliability Engineer (SRE) on a 24x7 frontline batch team acts as the primary guardian of critical data pipelines and scheduled workloads. Operating around the clock, these engineers actively monitor high-volume processing jobs, rapidly triaging alerts and executing tactical interventions to ensure critical batch processes meet strict business deadlines. When job failures or queue bottlenecks occur, they rely on operational runbooks and scripting to restore service, preventing downstream data delays for clients and internal systems. Beyond immediate troubleshooting, they continually automate repetitive operational tasks, tune monitoring alerts, and lead post-incident analysis to identify underlying root causes and build long-term pipeline resilience.

In This Role You’ll
  • Run and monitor scheduled batch jobs
  • Troubleshot system errors and performance issues
  • Do incident management to recover incidents and collaborate with application owners on problem management to understand the root cause and avoid reoccurrence
  • Maintain documentation and logs and work to improve processes (feedback loops)
  • Track metrics and KPIs on batch support activities
  • Coordinate with other departments to ensure job dependencies are met.
  • Automate routine tasks and deployments (as needed) using scripting and configuration management tools (e.g., Bash, Python, Ansible, Terraform)
About The Ideal Candidate
  • An ideal candidate for a 24x7 frontline batch SRE brings a distinct mix of operational calm, strong troubleshooting skills, communication, and an automation-first mindset.
  • Hands-On Troubleshooting: Demonstrated experience in production support or operations—ideally on a 24x7 rotation—where they’ve managed real-time incident response for Cloud environments (Azure)
  • Batch Processing Expertise: Familiarity with batch processing (ex: Control-M -BMC, SFTP), understanding how batches work and how to analyze logs, recover failing jobs, …
  • Observability & Triage: Experience using monitoring tools (Prometheus, Grafana, Splunk) for investigation
  • Strong Scripting & Systems Fundamentals: Proficiency in Python to write quick automation to optimize processes.
  • Grace Under Pressure: Keeps a cool head during high‑severity incidents, prioritizing rapid recovery and making sound judgment calls under strict deadlines.
  • Operational Discipline: Clear and concise communicator during shift handovers and active incidents, leaving detailed audit trails and updating runbooks so the rest of the team succeeds.
  • Toil-Averse Mindset: Inherently lazy about manual tasks in the best way possible—they get annoyed by repeating the same fix twice and prefer to automate it away permanently.
Required Qualifications
  • Bachelor’s degree in computer science, Engineering, or related field, or equivalent practical experience.
  • Basic understanding of Linux/Unix/Windows systems and networking.
  • Familiarity with cloud platforms
  • Basic understanding of batch processing (ex: Control-M -BMC, SFTP)
  • Familiarity with batch scripting and system architecture.
  • Ability to work independently and under pressure
What We Can Offer You
  • Get rewarded with competitive remuneration, individual and company annual bonus, vacation and holiday paid time off, health insurances and other competitive benefits.
  • Hybrid working model.
  • Professional development to broaden yourknowledge and enhance your skillswith on-line learning hubs packed with technical and soft skills training that allow you to develop and grow.
  • Enter a diverse and inclusive workplace, join one of the world’s top travel technology companies and take on a role that impacts millions of travelers around the globe.
Working at Amadeus, you will find
  • A critical mission and purpose - At Amadeus, you will be powering the future of travel and pursuing a critical mission and extraordinary purpose.
  • A truly global DNA - Everything at Amadeus is global, from our people to our business, which translates into our footprint, processes, and culture.
  • Great opportunities to learn - Learning happens all the time and in many ways at Amadeus, through on-the-job training, formal learning activities, and day-to-day interactions with colleagues.
  • A caring environment - Amadeus fosters a caring environment, nurturing both a fulfilling career and personal and family life. We care about our employees and strive to provide a supportive work environment.
  • A complete rewards offer - Amadeus provides attractive remuneration packages, covering all essential components of a competitive reward offer, including salary, bonus, equity, and benefits.
  • A diverse and inclusive community - We are committed to leveraging our uniquely diverse population to drive innovation, creativity, and collaboration across our organization.
  • A Reliable Company - Trust and reliability are fundamental values that drive our actions and shape long-lasting relationships with your customers, partners, and employees.
Diversity & Inclusion

Amadeus aspires to be a leader in Diversity and Inclusion in the tech industry, enabling every employee to reach their full potential by fostering a culture of belonging and fair treatment, attracting the best talent from all backgrounds, and as a role model for an inclusive employee experience.

Amadeus is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to gender, race, ethnicity, sexual orientation,age, beliefs, disability or any other characteristics protected by law.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Service Reliability Engineer
Service Reliability Engineer

Amadeus Hospitality • Bogotá

Híbrido
COP 90.000.000 - 150.000.000
Hybrid work model
Health insurance
Annual bonus
+1
Service Reliability Engineer
Service Reliability Engineer

Amadeus • Colombia

Presencial
COP 258.923.000 - 332.902.000
Competitive remuneration
Annual bonuses
Health insurance
+2
Senior Service Reliability Engineer
Senior Service Reliability Engineer

Amadeus • Bogotá

Híbrido
COP 90.000.000 - 150.000.000
Hybrid work in Bogota
Annual bonus
Health insurance
+1
Global Operations SRE
Global Operations SRE

1083 Amadeus IT Group Colombia, S.A.S. • Bogotá

Híbrido
COP 120.000.000 - 180.000.000
Hybrid working model
Health insurance
Annual bonus
+1
Global Operations SRE
Global Operations SRE

Amadeus Hospitality • Bogotá

Híbrido
COP 120.000.000 - 240.000.000
Senior Service Reliability Engineer
Senior Service Reliability Engineer

Amadeus Hospitality • Bogotá

Híbrido
COP 60.000.000 - 90.000.000
Annual bonus
Health insurance
Paid time off
+1
Senior Service Reliability Engineer
Senior Service Reliability Engineer

1083 Amadeus IT Group Colombia, S.A.S. • Bogotá

Híbrido
COP 180.000.000 - 320.000.000
Annual bonus
Health insurance
Software Development Engineers (All Levels)
Software Development Engineers (All Levels)

Amadeus • Colombia

Híbrido
COP 90.000.000 - 150.000.000
Competitive remuneration
Individual and company annual bonus
Paid vacation and holidays
+2
System Administrator
System Administrator

Amadeus Hospitality • Bogotá

Híbrido
COP 60.000.000 - 90.000.000
Competitive remuneration
Annual bonus
Paid time off
+3
Team Lead Engineering
Team Lead Engineering

Amadeus Hospitality • Bogotá

Híbrido
COP 120.000.000 - 240.000.000
Hybrid working model
Competitive remuneration with bonus
Learning & development opportunities
+1