Site Reliability Engineer - Argentina

teladoc

Argentina

Presencial

ARS 1.200.000 - 1.800.000

Jornada completa

Hace 7 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Teladoc Health is seeking a Site Reliability Engineer focused on Azure observability and reliability. You will design and implement monitoring, dashboards, and automated alerts across hybrid cloud environments to ensure service availability and performance.

The role emphasizes blameless post-incident reviews and strong collaboration with security and engineering teams. Responsibilities include defining SLI/SLOs, driving incident response playbooks, and contributing to disaster recovery planning.

Formación

  • Experience with Azure-based cloud platforms and monitoring.
  • Hands-on with enterprise observability tools (Datadog, Dynatrace, Elastic, Grafana, Prometheus).
  • Strong incident management and post-incident retro analysis.

Responsabilidades

  • Design, implement, and maintain observability across cloud environments and multi-cloud platforms.
  • Define SLIs/SLOs/SLAs to measure service health and customer experience.
  • Develop dashboards and automated alerting to identify degradations early.
  • Build on-call runbooks and post-incident retrospectives to drive improvements.
  • Collaborate with security, network, and engineering teams for compliant reliability.

Conocimientos

Azure expertise
Observability mindset
Incident management
Collaboration

Herramientas

Datadog
Dynatrace
Elastic
Grafana
Prometheus
LogicMonitor
Azure Monitor

Descripción del empleo

Summary of Position

We are seeking a highly skilled Site Reliability Engineer (SRE) with deep experience in Azure environments, specializing in Observability, Monitoring, and Incident Response. This role is critical to ensuring the availability, reliability, and performance of our hybrid cloud infrastructure and services. The ideal candidate will design and implement observability frameworks, drive automation in monitoring and alerting, and lead effective incident management processes across multi-cloud environments.

This position requires strong technical acumen in cloud-native operations, a proactive mindset toward reliability engineering, and the ability to collaborate with engineering, operations, and security teams to maintain mission-critical healthcare and enterprise workloads.

Essential Duties and Responsibilities
Observability & Monitoring
  • Design, implement, and maintain observability solutions across Azure (e.g., Azure Monitor, Datadog, Grafana-Prometheus, Dynatrace, Elastic).
  • Define and standardize SLIs/SLOs/SLAs to measure service health and customer experience.
  • Develop dashboards and automated alerting to proactively identify service degradations.
Incident Response
  • Build and maintain on-call runbooks and playbooks to reduce time-to-resolution.
  • Drive post-incident "blameless" retrospectives and continuous improvement initiatives.
Reliability Engineering
  • Develop automation for self-healing systems, monitoring remediation, and incident mitigation.
  • Contribute to disaster recovery and business continuity planning across multi-cloud platforms.
  • Work with engineering teams to design for resiliency, scalability, and reliability from the ground up.
Collaboration & Governance
  • Partner with security, network, and system engineering teams to ensure observability integrates with compliance and governance frameworks.
  • Advocate for best practices in cloud-native reliability engineering.
  • Mentor engineering staff in observability tools, monitoring strategies, and incident management.

The time spent on each responsibility reflects an estimate and is subject to change dependent on business needs.

Supervisory Responsibilities

No

Required Qualifications
  • Cloud Platforms: Expertise in Azure (VMs, AKS, Application Insights, Azure Monitor).
  • Observability Tools: Hands-on experience with enterprise observability platform such as Datadog and Dynatrace, Elastic, Grafana, Prometheus orLogicMonitor.
  • Monitoring & Alerting: Deep understanding of metrics, logs, traces, and distributed system monitoring.
Preferred Qualifications
  • Automation & Infrastructure as Code (IaC): Proficiency with Terraform, Bicep, Ansible, or similar tools to automate monitoring and remediation workflows.
  • Kubernetes Observability: Knowledge of AKS logging, tracing, and monitoring in containerized environments.
  • AI & Observability: Exposure to AI-driven monitoring, anomaly detection, or predictive alerting tools.
  • Programming/Scripting: Scripting skills in Python, PowerShell, or similar languages for automation and tool integration.
  • Chaos Engineering: Experience with resiliency testing and tools such as Gremlin or Chaos Mesh.
  • Healthcare & Compliance: Experience in healthcare IT environments with HIPAA, HITRUST, or other compliance frameworks.

As part of our hiring process, we verify identity and credentials, conduct interviews (live or video), and screen for fraud or misrepresentation. Applicants who falsify information will be disqualified.

Why join Teladoc Health?

Teladoc Health is transforming how better health happens. Learn how when you join us in pursuit of our impactful mission .

Chart your career path with meaningful opportunities that empower you to grow, lead, and make a difference.

Join a multi-faceted community that celebrates each colleague's unique perspective and is focused on continually improving, each and every day.

Contribute to an innovative culture where fresh ideas are valued as we increase access to care in new ways.

Enjoy an inclusive benefits program centered around you and your family, with tailored programs that address your unique needs.

Explore candidate resources with tips and tricks from Teladoc Health recruiters and learn more about our company culture by exploring #TeamTeladocHealth on LinkedIn .

As an Equal Opportunity Employer, we never have and never will discriminate against any job candidate or employee due to age, race, religion, color, ethnicity, nati

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí