Site Reliability Engineer II

Medallia

Tecámac

Híbrido

MXN 1.032.346 - 1.462.491

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Diversity and Inclusion Programs
Hybrid Work Model
Professional Development Opportunities

Descripción de la vacante

Medallia is seeking a Site Reliability Engineer II for their team in Tecámac. This role focuses on improving the reliability and scalability of their SaaS platform, pushing for automation in operations.

The ideal candidate will have over 2 years of experience in site reliability engineering or DevOps and proficiencies in Kubernetes and cloud services. The position operates in a hybrid model requiring some on-site work in the Mexico City area.

Formación

  • 2+ years of experience in Site Reliability Engineering, DevOps, or related roles.
  • Experience supporting production environments on Kubernetes or other container platforms.
  • Professional proficiency in written and spoken English.

Responsabilidades

  • Collaborate with engineering teams to enhance application reliability.
  • Operate production services in Kubernetes environments.
  • Build automation to reduce operational overhead.
  • Support CI/CD and GitOps-based deployment workflows.
  • Participate in incident response and operational improvements.

Conocimientos

Kubernetes
Cloud Infrastructure (AWS, GCP, OCI)
Linux Systems
Scripting (Python, Bash, Go)
CI/CD Pipelines
Networking Fundamentals
Troubleshooting Distributed Systems
On-call Support

Herramientas

GitOps
Terraform
Prometheus
Grafana

Descripción del empleo

Overview

Medallia is the pioneer and market leader in Experience Management. Our award‑winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees, patients, and residents alike.

We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company and greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees.

We empower exceptional people to create extraordinary experiences together. Bring your whole self.

The Role and Team

The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform. As an SRE II, you will help operate and improve the reliability, scalability, and performance of services running across Kubernetes‑based environments in cloud and hybrid infrastructure. You will work closely with software engineering teams to build automation, improve operational excellence, and support production services used globally by Medallia customers.

We are looking for engineers who enjoy solving complex technical problems, automating repetitive tasks, improving system reliability, and learning modern cloud‑native technologies in a fast‑paced environment. We value engineers who actively seek opportunities to improve scalability and operational efficiency through automation, AI‑assisted engineering workflows, and continuous process improvement. This role participates in a rotating on‑call schedule supporting production systems and services.

Engineering Leverage

At Medallia, we hire engineers who scale systems, teams, and outcomes through automation, platform thinking, and AI‑assisted engineering. We value engineers who challenge manual processes, reduce operational toil, and create reusable solutions that improve reliability and productivity for the broader engineering organization. Successful engineers do not simply solve problems—they eliminate recurring problems through automation, simplification, and self‑service capabilities.

Responsibilities
  • Collaborate with software engineering teams to improve application reliability, scalability, and operational maturity.
  • Operate and support production services running in Kubernetes environments.
  • Troubleshoot and resolve infrastructure and application issues across the full technology stack.
  • Build automation and tooling to reduce operational overhead and eliminate manual work.
  • Leverage AI‑assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational toil.
  • Identify opportunities to streamline operational processes through automation, AI‑enabled workflows, and self‑service solutions.
  • Create reusable solutions, tooling, and operational improvements that increase engineering leverage across the team.
  • Support CI/CD and GitOps‑based deployment workflows.
  • Develop and maintain infrastructure‑as‑code configurations and operational tooling.
  • Monitor system health, availability, and performance using observability and alerting platforms.
  • Participate in incident response, root cause analysis, and operational improvements.
  • Continuously improve reliability, deployment processes, and operational standards.

Candidates based in the Mexico City vicinity will be prioritized as this role is Hybrid, 3 days per week onsite.

Minimum Qualifications
  • 2+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Operations, or related roles.
  • Demonstrated experience supporting production environments running on Kubernetes or other containerized platforms.
  • Demonstrated experience with cloud infrastructure platforms such as AWS, OCI, or GCP.
  • Demonstrated experience with Linux systems administration and troubleshooting.
  • Demonstrated experience with scripting or programming languages such as Python, Bash, or Go.
  • Familiarity with CI/CD pipelines and Git‑based workflows.
  • Demonstrated understanding of networking fundamentals including DNS, load balancing, TLS/SSL, and routing concepts.
  • Demonstrated experience troubleshooting distributed systems and production incidents.
  • Ability to participate in an on‑call rotation supporting production systems.
  • Professional working proficiency in written and spoken English.
Preferred Qualifications
  • Experience with GitOps and tools such as ArgoCD.
  • Experience with infrastructure‑as‑code tools such as Terraform.
  • Familiarity with observability platforms such as Prometheus, Grafana, Loki, or OpenTelemetry.
  • Experience operating services in hybrid‑cloud or multi‑region environments.
  • Understanding of release strategies such as rolling deployments, canary releases, or blue/green deployments.
  • Familiarity with incident management and operational best practices.
  • Exposure to security and compliance concepts in production environments.
  • Experience using AI‑assisted development, automation, or operational tooling to improve engineering productivity and service reliability.
  • Demonstrated passion for automation, process improvement, and operational efficiency.
  • Strong communication and collaboration skills.

At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at ApplicantAccessibility@medallia.com. For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Software Engineering, Senior (Java, Kubernetes, Kafka)
Software Engineering, Senior (Java, Kubernetes, Kafka)

Medallia • Tecámac

Híbrido
MXN 1.250.000 - 1.608.000
Site Reilability Engineer
Site Reilability Engineer

Hcltech • Ecatepec de Morelos

Presencial
MXN 520.000 - 760.000
Life insurance
Major Medical Expenses Insurance
Minor Medical Expense Insurance
+4
SRE II: Automation, Kubernetes & Cloud Reliability
SRE II: Automation, Kubernetes & Cloud Reliability

Medallia • Tecámac

Híbrido
MXN 1.032.000 - 1.463.000
Diversity and Inclusion Programs
Hybrid Work Model
Professional Development Opportunities
Senior Talent Acquisition Partner, AI & Tech (LATAM)
Senior Talent Acquisition Partner, AI & Tech (LATAM)

Medallia • Tecámac

Híbrido
MXN 900.000 - 1.500.000
Senior Talent Acquisition Partner, AI & Tech (LATAM)
Senior Talent Acquisition Partner, AI & Tech (LATAM)

Medallia • Ciudad de México

Híbrido
MXN 2.092.000 - 3.140.000
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Rosarito

Híbrido
MXN 870.000 - 1.306.000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer ID60188
Site Reliability Engineer ID60188

AgileEngine • Ciudad de México

Híbrido
MXN 1.049.000 - 1.400.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Rosarito

A distancia
MXN 1.531.000 - 2.212.000
Mentorship and TechTalks
Competitive USD-based compensation
Work on modern solutions
+1
Site Reliability Engineer
Site Reliability Engineer

CTC • Estado de México

A distancia
MXN 1.433.000 - 1.793.000
Site Reliability Engineer (Remote)
Site Reliability Engineer (Remote)

itD • Estado de México

A distancia
MXN 1.438.000 - 1.978.000
Comprehensive medical benefits
401K with matching
Paid holidays
+1