Site Reliability Engineering, Engineer II

Medallia

Buenos Aires

Híbrido

ARS 2.400.000 - 3.600.000

Jornada completa

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca en este puesto — crea un currículum adaptado y una carta de presentación en aproximadamente un minuto.

Supera los filtros ATS

Descripción de la vacante

Medallia is seeking an SRE II to join its Site Reliability Engineering team. The role focuses on operating and improving reliability, scalability, and performance of services across Kubernetes-based cloud and hybrid environments.

Applicants will collaborate with software teams, participate in on-call rotations, and leverage automation and AI-assisted workflows to improve operational efficiency. Hybrid work requires three days onsite in the Buenos Aires area.

Formación

  • 2+ years in Site Reliability Engineering, DevOps, or related roles.
  • Experience supporting production environments on Kubernetes or similar.
  • Experience with AWS, OCI, or GCP cloud platforms.
  • Strong Linux systems administration and troubleshooting skills.
  • Experience with Python, Bash, or Go scripting.
  • Familiarity with CI/CD and Git workflows.
  • Understanding of DNS, load balancing, TLS/SSL, and routing.
  • Experience with distributed systems and incident handling.
  • Ability to participate in an on-call rotation.
  • Fluency in English, both speaking and writing.

Responsabilidades

  • Collaborate with software teams to improve reliability and scalability.
  • Operate and support production services in Kubernetes environments.
  • Troubleshoot infrastructure and application issues across the stack.
  • Build automation to reduce manual work and overhead.
  • Leverage AI-assisted tooling to speed troubleshooting and productivity.
  • Identify automation opportunities with AI-enabled workflows.
  • Create reusable tooling and improvements for the team.
  • Support CI/CD and GitOps deployment workflows.
  • Develop and maintain IaC configurations and tooling.
  • Monitor system health and performance with observability tools.
  • Participate in incident response and post-mortems.
  • Continuously improve reliability and deployment processes.

Conocimientos

SRE/DevOps
Kubernetes
Cloud platforms
Linux
Scripting: Python/Bash/Go
CI/CD pipelines
Networking basics
Distributed systems
On-call readiness
English fluency

Herramientas

ArgoCD
Terraform
Prometheus
Grafana
Loki
OpenTelemetry
Git

Descripción del empleo

Overview

Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud, leads the market in the management of experiences, insights, and actions for candidates, customers, employees, patients, and residents alike.


We believe that every experience is a memory that can last a lifetime. Experiences shape the way people feel about a company. And they greatly influence how likely people are to advocate, contribute, and stay. At Medallia, we are committed to creating a world where organizations are loved by their customers and their employees.


We empower exceptional people to create extraordinary experiences together.


Bring your whole self.


The Role and Team

The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform.

As an SRE II, you will help operate and improve the reliability, scalability, and performance of services running across Kubernetes-based environments in cloud and hybrid infrastructure. You will work closely with software engineering teams to build automation, improve operational excellence, and support production services used globally by Medallia customers.

We are looking for engineers who enjoy solving complex technical problems, automating repetitive tasks, improving system reliability, and learning modern cloud-native technologies in a fast-paced environment.

We value engineers who actively seek opportunities to improve scalability and operational efficiency through automation, AI-assisted engineering workflows, and continuous process improvement.

Please note this role participates in a rotating on-call schedule supporting production systems and services.

Responsibilities
  • Collaborate with software engineering teams to improve application reliability, scalability, and operational maturity.
  • Operate and support production services running in Kubernetes environments.
  • Troubleshoot and resolve infrastructure and application issues across the full technology stack.
  • Build automation and tooling to reduce operational overhead and eliminate manual work.
  • Leverage AI-assisted engineering tools and automation platforms to accelerate troubleshooting, improve productivity, and reduce operational toil.
  • Identify opportunities to streamline operational processes through automation, AI-enabled workflows, and self-service solutions.
  • Create reusable solutions, tooling, and operational improvements that increase engineering leverage across the team.
  • Support CI/CD and GitOps-based deployment workflows.
  • Develop and maintain infrastructure-as-code configurations and operational tooling.
  • Monitor system health, availability, and performance using observability and alerting platforms.
  • Participate in incident response, root cause analysis, and operational improvements.
  • Continuously improve reliability, deployment processes, and operational standards.

Candidates based in the Buenos Aires vicinity will be prioritized as this role is Hybrid, 3 days per week onsite.

Qualifications

Minimum Qualifications

  • 2+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Cloud Operations, or related roles.
  • Demonstrated experience supporting production environments running on Kubernetes or other containerized platforms.
  • Demonstrated experience with cloud infrastructure platforms such as AWS, OCI, or GCP.
  • Demonstrated experience with Linux systems administration and troubleshooting.
  • Demonstrated experience with scripting or programming languages such as Python, Bash, or Go.
  • Familiarity with CI/CD pipelines and Git-based workflows.
  • Demonstrated understanding of networking fundamentals including DNS, load balancing, TLS/SSL, and routing concepts.
  • Demonstrated experience troubleshooting distributed systems and production incidents.
  • Ability to participate in an on-call rotation supporting production systems.
  • Fluency in English, both oral and written.

Preferred Qualifications

  • Experience with GitOps and tools such as ArgoCD.
  • Experience with infrastructure-as-code tools such as Terraform.
  • Familiarity with observability platforms such as Prometheus, Grafana, Loki, or OpenTelemetry.
  • Experience operating services in hybrid-cloud or multi-region environments.
  • Understanding of release strategies such as rolling deployments, canary releases, or blue/green deployments.
  • Familiarity with incident management and operational best practices.
  • Exposure to security and compliance concepts in production environments.
  • Experience using AI-assisted development, automation, or operational tooling to improve engineering productivity and service reliability.
  • Demonstrated passion for automation, process improvement, and operational efficiency.
  • Strong communication and collaboration skills.

At Medallia, we celebrate diversity and recognize the value it brings to our customers and employees. Medallia is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age (40 and over), disability, genetic information, veteran status or military service, or any other status protected by state or local law. Individuals with a disability who need an accommodation to apply please contact us at ApplicantAccessibility@medallia.com. For information regarding how Medallia collects and uses personal information, please review our Privacy Policies. Applications will be accepted for 30 days from the date this role was posted or until the role has been filled.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior Data Platform Engineer
Senior Data Platform Engineer

Medallia • Municipio de Esquel

Presencial
ARS 121.043.000 - 196.696.000
Senior Data Platform Engineer
Senior Data Platform Engineer

Medallia • Buenos Aires

Presencial
ARS 136.275.000 - 211.983.000
Staff Data Platform Engineer
Staff Data Platform Engineer

Medallia • Municipio de Esquel

Presencial
ARS 211.983.000 - 302.833.000
Senior Technical Program Manager, AI
Senior Technical Program Manager, AI

Medallia • Municipio de Esquel

Presencial
ARS 209.782.000 - 314.673.000
SRE II: Cloud-Native Reliability & Automation Engineer
SRE II: Cloud-Native Reliability & Automation Engineer

Medallia • Buenos Aires

Híbrido
ARS 2.400.000 - 3.600.000
Senior Platform Software Engineer (Streaming and Event-Driven Platforms)
Senior Platform Software Engineer (Streaming and Event-Driven Platforms)

Medallia • Municipio de Esquel

Presencial
ARS 121.133.000 - 181.700.000
Software Engineer II, Text Analytics (Java, NLP)
Software Engineer II, Text Analytics (Java, NLP)

Medallia • Buenos Aires

Híbrido
ARS 1.200.000 - 1.800.000
Senior Technical Program Manager, AI
Senior Technical Program Manager, AI

Medallia • Buenos Aires

Presencial
ARS 179.813.000 - 269.719.000
Senior IT Systems Engineer, Collaboration Tools (Google Workspace)
Senior IT Systems Engineer, Collaboration Tools (Google Workspace)

Medallia • Buenos Aires

Híbrido
ARS 1.200.000 - 1.800.000
Product Security Engineer II
Product Security Engineer II

Medallia • Buenos Aires

Presencial
ARS 52.104.000 - 81.878.000