Lead Site Reliability Engineer

Capital One National Association

Ciudad de México

Presencial

MXN 1.200.000 - 2.000.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Transforma esta oferta en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.

Supera los filtros ATS

Descripción de la vacante

Capital One Technology Labs Mexico is building a Site Reliability Engineering center in Mexico City and is hiring a Manager-level Backend Engineer to own the reliability of settlement platforms. You will work across on-prem data centers and AWS, alongside UK engineers, to automate and observe critical batch processes.

This foundational role requires strong SRE experience, cloud-native skills, and proficiency in Java or Python, with a focus on reducing toil and ensuring regulatory readiness for

Formación

  • Bachelor's degree required.
  • 6+ years in SRE, production operations, or reliability engineering.
  • 5+ years in Java, Python, or Go.
  • 4+ years with cloud-native technologies (AWS/Azure/GCP).
  • 3+ years with container orchestration (Docker/Kubernetes).
  • Proficient in Shell/Bash scripting and Unix/Linux.

Responsabilidades

  • Own reliability for batch settlement systems and ensure cycle windows are met.
  • Build observability dashboards, alerts, and anomaly detection.
  • Automate operational toil such as certificate rotation and provisioning.
  • Collaborate with UK engineers on compliance windows and SLA adherence.
  • Participate in incident management and root cause analysis.
  • Contribute to audit-ready artifacts for SOX and PCI-DSS.

Conocimientos

SRE
DevOps
Java
Python
Go
Cloud Native (AWS/Azure/GCP)
Docker/Kubernetes
Shell scripting
Unix/Linux

Educación

Bachelor's degree

Herramientas

Docker
Kubernetes
HashiCorp Vault
Datadog
OpenShift
AWS

Descripción del empleo

WeWork Reforma Latino (97001), Mexico, Ciudad de Mexico, Ciudad de MexicoLead Site Reliability Engineer

We're building a Site Reliability Engineering center in Mexico City, and we're hiring a Manager-level Backend Engineer to own the reliability and operational maturity of our settlement platforms. These are batch-critical systems that process every credit and debit transaction across the network.

This is a foundational role. You'll be one of the first engineers in CDMX responsible for ensuring settlement cycles complete accurately, on time, and in compliance with SOX and PCI-DSS requirements. You'll work across hybrid infrastructure (on-prem data centers and AWS), partner closely with UK-based engineers, and build the automation and observability that allows Mexico City to operate settlement.

What You'll Do
  • Own reliability for batch settlement systems - ensure cycle completion windows are met, data integrity is maintained, and failures are detected before they reach downstream consumers
  • Build and improve observability for settlement pipelines - dashboards, alerts, and anomaly detection that make system health legible and reduce reliance on tribal knowledge
  • Drive automation of operational toil - certificate rotation, environment provisioning, compliance artifact generation, and manual validation steps that currently require human intervention
  • Partner with UK-based settlement engineers - acquire domain expertise on Durbin compliance windows, cross-border DCI routing, and acquirer/issuer SLA adherence
  • Participate in incident management - respond to settlement failures, drive root cause analysis, and implement durable fixes that prevent recurrence
  • Contribute to regulatory readiness - ensure SRE practices produce audit-ready artifacts for SOX and PCI-DSS exams without manual toil
What Success Looks Like
  • Independently validate and troubleshoot settlement cycle failures
  • At least two manual settlement operations processes fully automated
  • Settlement observability coverage sufficient to detect anomalies before cycle deadlines
  • Documented runbooks and severity criteria for all critical settlement failure modes
The Environment

You'll work with batch processing systems that handle financial transactions across multiple on-prem data centers with active/active and active/passive configurations. The stack includes Java, Python, shell scripting, SQL, AWS, Kubernetes, OpenShift containers, Datadog, Observe, and legacy payment platforms. CI/CD pipelines, API automation, and secret management via HashiCorp Vault are part of daily operations. You'll leverage agentic AI automation (Claude Code or others) to accelerate development and build automation solutions. You'll need strong troubleshooting and debugging skills and be comfortable with both modern cloud-native tooling and traditional enterprise batch systems.

Basic Qualifications
  • Professional English fluency
  • Bachelor's degree
  • At least 6 years of experience in SRE, production operations, or reliability engineering
  • Experience in DevOps Engineering (internship experience does not apply)
  • 5+ years of experience in at least one of the following: Java, Python, Go
  • At least 4 years of experience with Cloud Native technologies (Amazon Web Services, Microsoft Azure, Google Cloud Platform)
  • 3+ years of experience with container orchestration services including Docker or Kubernetes
  • Experience with Shell or Bash scripting
  • At least 3 years of Unix or Linux system administration experience
Preferred Qualifications
  • Experience developing automation solutions using agentic AI tools (Claude Code, Copilot CLI)
  • Troubleshooting and debugging skills across distributed systems
  • Familiarity with payments, financial services, or other regulated high-availability domains
  • Knowledge or experience of Networking concepts (TCP/DNS/TLS)

At Capital One, we respect individual differences in culture, religion, and ethnicity. Likewise, we promote equal opportunities and development for all personnel. In the hiring process, we seek to provide equal employment opportunities to candidates, regardless of race, color, religion, gender, sexual orientation, marital or civil status, national origin, disability, or any other situation protected by federal, state, or local laws.

Capital One does not provide, endorse nor guarantee and is not liable for third-party products, services, educational tools or other information available through this site.

Capital One Financial is made up of several different entities. Please note that any position posted in Canada is for Capital One Canada, any position posted in the United Kingdom is for Capital One Europe, any position posted in the Philippines is for Capital One Service Corp (COPSSC), and any position posted in Mexico is for Capital One Technology Labs Mexico.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Lead Site Reliability Engineer
Lead Site Reliability Engineer

Military Friendly Company • Ciudad de México

Híbrido
MXN 1.200.000 - 1.800.000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Capital One Group • Ciudad de México

Híbrido
MXN 1.200.000 - 1.800.000
Sr. Manager SRE (Individual Contributor)
Sr. Manager SRE (Individual Contributor)

Capital One • Ciudad de México

Presencial
MXN 1.400.000 - 2.100.000
Senior Site Reliability Engineer – Settlement Platforms
Senior Site Reliability Engineer – Settlement Platforms

Capital One National Association • Ciudad de México

Presencial
MXN 1.200.000 - 2.000.000
Senior SRE Lead - Settlement Systems Reliability
Senior SRE Lead - Settlement Systems Reliability

Capital One • Ciudad de México

Presencial
MXN 2.100.472 - 2.625.591
Senior Site Reliability Engineer - Batch Observability
Senior Site Reliability Engineer - Batch Observability

Military Friendly Company • Ciudad de México

Híbrido
MXN 1.200.000 - 1.800.000
Senior Director, Software Engineering
Senior Director, Software Engineering

Capital One National Association • Ciudad de México

Presencial
MXN 1.800.000 - 2.400.000
Senior SRE - Batch Settlement & Observability
Senior SRE - Batch Settlement & Observability

Capital One Group • Ciudad de México

Híbrido
MXN 1.200.000 - 1.800.000
Senior Manager, Software Engineering - Full Stack (People Manager)
Senior Manager, Software Engineering - Full Stack (People Manager)

Capital One • Ciudad de México

Presencial
MXN 1.500.000 - 2.300.000
Distinguished Engineer
Distinguished Engineer

Capital One Group • Ciudad de México

Híbrido
MXN 900.000 - 1.500.000