Senior Site Reliability & Observability Engineer (SRE)

TransUnion

San Miguel del Resgate

Presencial

MXN 600.000 - 900.000

Jornada completa

14 días+
Generador de candidaturas

No envíes un currículum genérico: crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Descripción de la vacante

TransUnion in Mexico is seeking a Consultant, IT Support to join the Reliability Engineering team on-site. You will uphold the reliability, stability, and performance of mission-critical platforms using SRE best practices and robust incident management.

You will monitor observability data, automate toil, and collaborate with infrastructure, development, and security teams to reduce outages and improve deployment readiness.

Formación

  • Experience in Site Reliability Engineering (SRE) or related fields.
  • Strong knowledge of distributed systems, observability, and incident management.
  • Familiarity with SLOs, SLIs, and error budgets.

Responsabilidades

  • Ensure the reliability, stability, and availability of mission-critical services through SRE best practices.
  • Operate and improve the observability platform including metrics, logs, traces, and alerts.
  • Monitor critical systems and respond to incidents to minimize disruption.
  • Lead major incident response, coordinating recovery and stakeholder communication.
  • Conduct root cause analysis and blameless post-mortems to drive improvements.

Conocimientos

SRE
Incident management
Observability
Kubernetes
Cassandra
Kafka

Educación

Bachelor's degree in Computer Science, Systems Engineering, Telecommunications Engineering, or related

Herramientas

Apache Cassandra
Kafka
Kubernetes
Relational databases
NoSQL platforms

Descripción del empleo

TransUnion's Job Applicant Privacy Notice

Team Overview

The Reliability Engineering team ensures the stability, availability, and performance of Buró de Crédito’s mission-critical platforms and services. Through observability, automation, and incident management practices, the team drives operational excellence and continuous improvement. Working closely with Infrastructure, Development, Database, and Security teams, they help maintain resilient systems that support critical business operations. This job is assigned as On-Site Essential and requires in-person work at an assigned TU office location as a condition of employment.

Role Overview And Core Responsibilities
  • Ensure the reliability, stability, and availability of mission-critical services through Site Reliability Engineering (SRE) best practices.
  • Operate and continuously improve the organization's observability platform, including metrics, logs, traces, and alerting capabilities.
  • Monitor critical systems proactively and respond to operational incidents to minimize service disruption and business impact.
  • Lead major incident response activities, coordinating recovery efforts and stakeholder communication during service outages.
  • Conduct root cause analysis (RCA) and facilitate blameless post-mortems to identify systemic improvements and prevent recurrence.
  • Define, monitor, and improve Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.
  • Automate operational processes and reduce manual effort (toil) through scripting, infrastructure automation, and platform engineering practices.
  • Support and troubleshoot distributed environments including Apache Cassandra, Kafka, Kubernetes, relational databases, and NoSQL platforms.
  • Collaborate with development and infrastructure teams to improve deployment reliability, operational readiness, and platform performance.
  • Drive continuous improvement initiatives focused on monitoring, availability, scalability, resilience, and operational efficiency.
Required Knowledge And Experiences
  • Bachelor's degree in Computer Science, Systems Engineering, Telecommunications Engineering, or a related technical discipline, providing the foundation required to manage complex distributed environments.
  • Proven experience in Site Reliability Engineering (SRE), Production Operations, Platform Engineering, Infrastructure Engineering, or Reliability-focused roles.
  • Strong knowledge of distributed systems, observability practices, incident management, and troubleshooting methodologies for mission-critical environments.
  • Experience managing major incidents, root cause analysis processes, service restoration activities, and operational excellence initiatives.
  • Understanding of reliability frameworks including SLOs, SLIs, error budgets, continuous improvement, and service management best practices.
TransUnion Overview

At TransUnion, we encourage and are committed to creating a real, positive impact and shared sense of purpose within our Workforce for Good, which empowers our people to grow, innovate and contribute to a better future for our communities and customers. We strive to build an environment where our associates are in the driver’s seat of their professional development— while having access to help along the way. We recognize that success comes when our associates thrive both professionally and personally; that’s why we prioritize work/life flexibility and offer resources for our teams across the globe to collaborate and drive excellence.

Be a part of our Workforce for Good – you’ll work with great people, pioneering products and cutting-edge technology.

TransUnion Job Title

Consultant, IT Support

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior SRE: Observability & Reliability Architect
Senior SRE: Observability & Reliability Architect

TransUnion • San Miguel del Resgate

Presencial
MXN 600.000 - 900.000
Sr. Analyst - Data Science
Sr. Analyst - Data Science

TransUnion • San Miguel del Resgate

Híbrido
MXN 360.000 - 600.000
Senior Consultant – Research and Consulting
Senior Consultant – Research and Consulting

TransUnion • San Miguel del Resgate

Híbrido
MXN 900.000 - 1.300.000
Senior Consultant – Research and Consulting
Senior Consultant – Research and Consulting

TransUnion • México

Presencial
MXN 900.000 - 1.300.000
Consultant Compliance
Consultant Compliance

TransUnion • San Miguel del Resgate

Híbrido
MXN 720.000 - 1.080.000
Sr Consultant, Talent Acquisition
Sr Consultant, Talent Acquisition

TransUnion • San Miguel del Resgate

Híbrido
MXN 900.000 - 1.300.000
Ingeniero Sr de Redes y Conectividad
Ingeniero Sr de Redes y Conectividad

TransUnion • San Miguel del Resgate

Presencial
MXN 900.000 - 1.300.000
Advisor, Fraud Solutions Consulting MX
Advisor, Fraud Solutions Consulting MX

TransUnion • San Miguel del Resgate

Híbrido
MXN 520.000 - 780.000
Site Reliability Engineer
Site Reliability Engineer

CTC • Estado de México

A distancia
MXN 1.433.000 - 1.793.000
Site Reliability Engineer
Site Reliability Engineer

Pyramid Consulting, Inc • Estado de México

Presencial