Site Reliability Engineer (SRE)

Infosys Limited

Santiago

Presencial

CLP 20.088.000 - 37.944.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Infosys is seeking a Site Reliability Engineer (SRE) to ensure the reliability, availability, and performance of digital services in production. You will balance stability with rapid delivery, strengthen operational resilience through automation, and collaborate with application and platform teams across the Chilean edge.

The role emphasizes proactive reliability practices and incident management within a POD-based delivery model.

Formación

  • Proven experience as SRE, Production Engineer, or similar role.
  • Strong background in production systems and reliability engineering.
  • Experience with Cloud platforms, CI/CD pipelines, and monitoring/observability tools.

Responsabilidades

  • Define, implement, and maintain SLIs and SLOs (availability, latency, error rates).
  • Lead or coordinate incident response (L2/L3) and drive blameless postmortems.
  • Design and implement automation for deployments, monitoring, health checks, and self-healing.
  • Support change tracking, rollback strategies, and production readiness before release.
  • Define operational KPIs and drive data-driven decisions to reduce toil.

Conocimientos

SRE Experience
Production Engineering
Cloud platforms
CI/CD pipelines
Monitoring tools
POD model
Problem solving

Descripción del empleo

Domain|Infrastructure-Information Security Management|Information Security Compliance

Domain

Delivery

Interest Group

Infy Chile

Company

IL Chile

Requisition ID

152099BR

Job Description – Site Reliability Engineer (SRE)
Role Purpose

The Site Reliability Engineer (SRE) is responsible for ensuring the reliability, availability, and performance of digital services in production, balancing service stability with the ability to deliver change at speed. The role focuses on strengthening operational resilience through engineering, automation, and proactive reliability practices, working closely with application and platform teams.

Scope of the Role

The locally applied SRE role covers:

  • Associated infrastructure (cloud, CI/CD pipelines, integrations)
  • Continuous operations (24/7 reliability mindset, not necessarily shift-based)
  • Production changes (deployments, configurations)
  • Incidents, problems, and service degradations
  • Continuous improvement of stability and operational efficiency
Job Description – Site Reliability Engineer (SRE)
Role Purpose

The Site Reliability Engineer (SRE) is responsible for ensuring the reliability, availability, and performance of digital services in production, balancing service stability with the ability to deliver change at speed. The role focuses on strengthening operational resilience through engineering, automation, and proactive reliability practices, working closely with application and platform teams.

Scope of the Role

The locally applied SRE role covers:

  • Production digital services (applications, platforms, data products)
  • Associated infrastructure (cloud, CI/CD pipelines, integrations)
  • Continuous operations (24/7 reliability mindset, not necessarily shift-based)
  • Production changes (deployments, configurations)
  • Incidents, problems, and service degradations
  • Continuous improvement of stability and operational efficiency
Key Responsibilities
Service Reliability & Availability
  • Define, implement, and maintain SLIs and SLOs (availability, latency, error rates)
  • Continuously monitor service health and anticipate degradations
  • Ensure services operate within business‑agreed reliability thresholds
  • Manage reliability trade‑offs between speed and stability
Incident & Problem Management
  • Lead or coordinate response to relevant incidents (L2/L3)
  • Ensure:
    • Rapid and structured diagnosis
    • Safe service restoration
    • Clear and effective communication
  • Facilitate blameless postmortems
  • Convert recurring incidents into engineering improvement backlog
  • Drive long‑term remediation rather than reactive firefighting
Automation & Operational Excellence
  • Identify repetitive and manual operational tasks
  • Design and implement automation for:
    • Deployments
    • Monitoring and alerting
    • Health checks
    • Basic recovery and self‑healing (where applicable)
  • Reduce toil and increase system resilience through engineering solutions
Change Governance & Production Readiness
  • Support vendor and internal team change tracking
  • Ensure changes:
    • Are traceable
    • Have defined rollback strategies
    • Minimize operational risk
  • Validate operational readiness before production
  • Participate early in solution and architecture design from a reliability perspective (early involvement)
Metrics, Observability & Continuous Improvement
  • Define and maintain near real‑time operational KPIs (“service pulse”)
  • Ensure every deviation has:
    • Clear ownership
    • Defined corrective actions
  • Prevent reactive operations by driving data‑driven decision making
  • Support identification, prioritization, and planning of technical debt remediation
What This Role Is Not
  • A dedicated incident operator only
  • An advanced Service Desk
  • The sole owner of service stability (reliability is shared)
  • A gatekeeper blocking changes without technical justification
  • The owner of contractual MOPs
  • A commercial or account management role
  • The customer‑side account or delivery lead
Experience & Profile (Indicative)
  • Proven experience as SRE, Production Engineer, or similar role
  • Strong background in production systems and reliability engineering
  • Experience working with:
    • Cloud platforms
    • CI/CD pipelines
    • Monitoring and observability tools
  • Comfortable operating in product‑oriented or POD‑based team models
  • Strong problem‑solving, communication, and collaboration skills
Operating Model Alignment
  • Works embedded or as an enabling function with PODs
  • Focused on enablement and reliability patterns, not centralized control
  • Promotes shared ownership of reliability
About Us

Infosys is a global leader in next‑generation digital services and consulting. We enable clients in more than 50 countries to navigate their digital transformation. With over four decades of experience in managing the systems and workings of global enterprises, we expertly steer our clients through their digital journey. We do it by enabling the enterprise with an AI‑powered core that helps prioritize the execution of change. We also empower the business with agile digital at scale to deliver unprecedented levels of performance and customer delight. Our always‑on learning agenda drives their continuous improvement through building and transferring digital skills, expertise, and ideas from our innovation ecosystem.

EEO

Infosys provides equal employment opportunities to applicants and employees without regard to race; color; sex; gender identity; sexual orientation; religious practices and observances; national origin; pregnancy, childbirth, or related medical conditions; or disability.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Site Reliability Engineer
Site Reliability Engineer

Infosys • Santiago

Presencial
CLP 24.000.000 - 42.000.000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Infosys • Santiago

Presencial
CLP 25.000.000 - 45.000.000
Site Reliability Engineer — Build Resilient, Automated Systems
Site Reliability Engineer — Build Resilient, Automated Systems

Infosys • Santiago

Presencial
CLP 24.000.000 - 42.000.000
Site Reliability Engineer - Automate & Stabilize Production
Site Reliability Engineer - Automate & Stabilize Production

Infosys Limited • Santiago

Híbrido
CLP 20.088.000 - 37.944.000
Operational Continuity Engineer
Operational Continuity Engineer

Infosys Limited • Santiago

Híbrido
CLP 7.000.000 - 11.000.000
Senior Network Engineer
Senior Network Engineer

Infosys Limited • Santiago

Híbrido
CLP 18.000.000 - 32.000.000
Site Reliability Engineer
Site Reliability Engineer

Grupo Falabella • Santiago

Presencial
CLP 33.480.000 - 60.264.000
DevOps / SRE Engineer
DevOps / SRE Engineer

IT NT Solutions • Chile

Presencial
CLP 60.000.000 - 90.000.000
Renta competitiva
Horarios flexible
Seguro de salud y dental
+1
SRE / Reliability Engineer — Remote (Chile)
SRE / Reliability Engineer — Remote (Chile)

Forest • Santiago

Híbrido
CLP 35.587.000 - 62.278.000
20 days of annual leave
Comprehensive Private Medical Insurance
£1000 personal development budget
+2
Ingeniero/a SRE Ssr - Nivel Ingles Avanzado
Ingeniero/a SRE Ssr - Nivel Ingles Avanzado

Haibu Solutions Spa • Santiago

Presencial
Día libre por cumpleaños
Seguro complementario de salud
Bono de vacaciones
+2