Lead Site Reliability Engineer - Cloud & AI Ops

JobCubby

Ciudad de México

Presencial

MXN 900.000 - 1.300.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Ventajas ofrecidas por este puesto de trabajo

Flexible work options
Wellness program
Discounts & scholarships
Professional development programs
Recognition programs

Descripción de la vacante

PepsiCo is seeking a Lead SRE Engineer to drive reliability across our global consumer, commercial, and supply chain applications. You will shape SRE practices, orchestrate across teams, and focus on proactive incident prevention and automation.

The role demands 8+ years in SRE and expertise in Python, cloud platforms, and monitoring tools. You will collaborate with engineering and support teams to ensure high availability and optimal performance for PepsiCo’s digital product portfolio.

Formación

  • 8+ years of work experience evolving to a SRE engineer with 3-5 years of experience in continuously improving and transforming IT operations ways of working.
  • Bachelor’s degree in Computer Science, Information Technology or a related field.
  • Proven experience as an SRE in designing the events diagnostics, performance measures and alert solutions to meet the SLA/SLO/SLIs.
  • The ideal Engineer will be highly quantitative, have great judgment, able to connect dots across ecosystems, and efficiently work cross-functionally across teams to ensure SRE orchestrating solutions are meeting customer/end-user expectations.
  • The candidate will take a pragmatic approach resolving incidents, including the ability to systemically triangulate root causes and work effectively with external and internal teams to meet objectives.
  • A strong expertise of SRE (Software Reliability Engineering) and IT Service Management (ITSM) processes with a track record for improving service offerings – pro-actively resolving incidents, providing a seamless customer/end-user experience and proactively identifying and mitigating areas of risk.
  • Hands on experience in Python, SQL /No-SQl( MySQL, Mongo DB, Cassandra, Postgress), AppDynamics, ELK Stack Grafana, Splunk, Dynatrace, Kafka and any SRE Ops toolsets.
  • A firm understanding of cloud archtitecture for distributed environments.
  • Front-end technologies: HTML, CSS, JavaScript, and frameworks like React, Angular, or Vue.js.
  • Back-end technologies: Server-side languages (Java, Spring Boot, and related technologies that build the server-side logic, APIs, and database interaction with MySQL, MongoDB, Cassandra, Couchbase)
  • Infrastructure: Azure/AWS cloud platforms and/or Client / server environments.
  • Prior experience involving in shaping transformation developing SRE solutions would be a plus.

Responsabilidades

  • Drive new shift left activities critical to apply Site Reliability Engineering (SRE) and quality assurance principles within the application design / Project roadmap that enablees resilient outcomes.
  • Apply pre-emptive approach into production minimizing business impact, via SRE-driven orchestration of connecting all components of the ecosystem diagnosing anomalies prior to user & remediating through automation.
  • This is a critical enabler achieving a high resiliency during operations and also continuously improving through design during the software development lifecycle.
  • The Lead SRE design & support engineer is integral part of the global team with its main purpose to provide a delightful customer experience for the user of the global consumer, commercial, supply chain and enablement functions in the PepsiCo digital products application portfolio of 260+ applications, enabling a full SRE Practice incident prevention / proactive resolution model.
  • The scope of this role is focussed on the cloud architecture application full stack devlopment, B2B pepsiconnect and Direct to Customer and other S&T roadmap applications. Ensures that PepsiCo DPA applications service performance, reliability and availability expected by our customers and internal groups.
  • It requires a blend of technical expertise on SRE tools, modern applications cloud architecture i.e. full stack, IT operations experience, and analytics & influence skills.
  • Please see responsibilities list in the source text for additional items.

Conocimientos

SRE Engineer
Python
Cloud Architecture
ITSM

Educación

Bachelors in CS or related

Herramientas

AppDynamics
ELK Stack
Grafana
Splunk
Dynatrace
Kafka
SRE Tools

Descripción del empleo

PepsiCo is seeking a Lead SRE Engineer to drive reliability across our global consumer, commercial, and supply chain applications. You will shape SRE practices, orchestrate across teams, and focus on proactive incident prevention and automation.

The role demands 8+ years in SRE and expertise in Python, cloud platforms, and monitoring tools. You will collaborate with engineering and support teams to ensure high availability and optimal performance for PepsiCo’s digital product portfolio.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Lead Cloud SRE: Reliability, Automation & Scale
Lead Cloud SRE: Reliability, Automation & Scale

EPAM Systems, Inc. • Morelia

Presencial
MXN 700.000 - 1.100.000
Healthcare benefits
Paid time off and sick leave
Upskilling and certification courses
+3
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

EPAM Systems, Inc. • Morelia

Presencial
MXN 700.000 - 1.100.000
Healthcare benefits
Paid time off and sick leave
Upskilling and certification courses
+3
Disaster Recovery Engineer - Cloud & Hybrid Resiliency
Disaster Recovery Engineer - Cloud & Hybrid Resiliency

PepsiCo • Ciudad de México

Presencial
MXN 600.000 - 850.000
Digital learning platforms
Leadership development programs
Flexible working arrangements
+1
Senior SRE: Cloud Infra, CI/CD & Resilient Systems
Senior SRE: Cloud Infra, CI/CD & Resilient Systems

EPAM Systems • México

Presencial
MXN 900.000 - 1.300.000
International projects
Global teams
Employee financial programs
+5
Senior Site Reliability Engineer - Scale, Automate, Resilient Ops
Senior Site Reliability Engineer - Scale, Automate, Resilient Ops

Mastercard • Ciudad de México

Presencial
MXN 1.200.000 - 2.000.000
Lead Platform Engineer
Lead Platform Engineer

PepsiCo • Ciudad de México

Híbrido
MXN 1.200.000 - 1.800.000
Opportunities for learning and development
Flexibility program for work-life balance
Recognition programs for achievements
+1
Site Reliability Engineer
Site Reliability Engineer

Tata Consultancy Services • Ciudad de México

Presencial
Senior SRE Lead (IC): Reliability & AI Automation
Senior SRE Lead (IC): Reliability & AI Automation

Hobbsnews • México

A distancia
MXN 1.750.000 - 2.276.000
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Rosarito

Híbrido
MXN 870.000 - 1.306.000
Professional growth
Competitive compensation
Exciting projects
+1
Remote Site Reliability Engineer — Cloud & Automation Leader
Remote Site Reliability Engineer — Cloud & Automation Leader

AgileEngine • Rosarito

A distancia
MXN 1.531.000 - 2.212.000
Mentorship and TechTalks
Competitive USD-based compensation
Work on modern solutions
+1