Site Reliability Engineer I

YUM

Plano (TX)

Hybrid

USD 96,000 - 120,000

Full time

4 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Vacation time
Paid holidays
Floating day
Half-day Fridays
Volunteer time
Medical insurance
401(k)

Job summary

KFC is seeking a Site Reliability Engineer I in Plano, Texas to support the reliability and operational health of production systems. You will work with Platform Engineering, Development, and Product teams to monitor, troubleshoot, and automate in a hybrid work setup.

The role emphasizes incident response, observability, cloud infrastructure, and IaC. A Bachelor’s degree is preferred along with 1–3+ years in SRE/DevOps and related technologies.

Qualifications

  • Bachelor’s degree or equivalent technical experience.
  • 1–3+ years in SRE/DevOps/platform engineering or related role.
  • Experience with cloud platforms (AWS/Azure/GCP).
  • Experience with monitoring/observability tools.
  • Familiarity with Linux and networking fundamentals.
  • Experience with containerized environments (Kubernetes/EKS).
  • Scripting or programming in Python/Go/Bash/REST APIs.
  • Experience with IaC tools (Terraform/Ansible/CloudFormation).
  • Familiarity with GitLab CI/CD.

Responsibilities

  • Monitor and troubleshoot production systems to improve reliability.
  • Operate within defined services and contribute to reliability improvements.
  • Build and maintain dashboards and alerts in observability tools.
  • Support cloud-based and containerized infrastructure.
  • Contribute to incident response and post-incident actions.
  • Develop automation to reduce toil and manual work.
  • Participate in 24/7 on-call rotation.
  • Document runbooks and perform RCA and corrective actions.

Skills

Troubleshooting
Analytical thinking
Communication
Collaboration
On-call support

Education

Bachelor’s degree or equivalent

Tools

Datadog
Elastic
Dynatrace
Prometheus
Grafana
Jira
Confluence
Incident.io
Kubernetes
EKS
Python
Go
Bash
REST APIs
Terraform
Ansible
CloudFormation
GitLab
CI/CD

Job description

It all started with one cook who created a finger lickin' good recipe more than 75 years ago, a list of secret herbs and spices scratched out on the back of the door to his kitchen.

Site Reliability Engineer I

KFC

Hybrid

  • The Site Reliability Engineer I supports the reliability, availability, and operational health of KFC production systems and applications. This role works with Platform Engineering, Development, and Product teams to monitor and troubleshoot production environments, improve observability, contribute to automation, and support cloud-based infrastructure.
  • The SRE I is expected to operate within defined systems and services, contribute to reliability improvements, participate in incident response, and build increasing technical ownership over time.
  • Observability – Monitor KFC production environments using tools such as Datadog. Create and improve monitors, dashboards, alerts, logging, and other telemetry used to proactively identify reliability and performance issues.
  • Incident Response – Participate in the investigation, troubleshooting, and restoration of production systems. Use metrics, logs, traces, and application telemetry to identify issues and support technical resolution and post-incident actions.
  • Cloud & Infrastructure – Support and improve cloud-based and containerized infrastructure, including compute, networking, application services, and supporting platform components.
  • Infrastructure as Code – Contribute to the deployment and maintenance of infrastructure using technologies such as Terraform, Ansible, and GitLab in partnership with SRE and Platform Engineering teams.
  • Automation – Develop basic scripts, tools, and automated workflows using technologies such as Python, Go, Bash, and REST APIs to reduce repetitive operational work and toil.
  • Reliability Engineering – Contribute to improvements in system availability, scalability, resilience, and performance by identifying operational gaps and implementing defined reliability improvements.
  • Deployment & Production Support – Assist with deployment readiness, production validation, change verification, and monitoring following application or infrastructure releases.
  • Documentation & Continuous Improvement – Maintain operational runbooks and contribute to root cause analysis, corrective actions, and improvements identified through incidents and service health reviews.
  • Participate in the SRE 24/7 on-call rotation to detect, respond to, and resolve production issues.
  • Education/Certifications – Bachelor’s Degree preferred or equivalent technical experience.
  • 1 to 3+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, infrastructure, application support, software engineering, or a related technical role.
  • 1+ year’s experience with cloud platforms such as AWS, Azure, or GCP, including compute and networking concepts.
  • 1+ year’s experience with monitoring and observability platforms such as Datadog, Elastic, Dynatrace, Prometheus, or Grafana.
  • 1+ year’s experience with Incident and Problem Management tools such as ServiceNow, Jira, Confluence, Incident.io.
  • Familiarity with Linux-based systems and command-line utilities.
  • Working knowledge of networking fundamentals, including DNS, load balancing, routing, and TLS.
  • Experience with containerized environments and orchestration tools such as Kubernetes or EKS.
  • Familiarity with scripting or programming languages such as Python, Go, Bash, or REST APIs.
  • Familiarity with Infrastructure as Code tools such as Terraform, Ansible, or CloudFormation.
  • Familiarity with CI/CD tools and practices, including GitLab.
  • Strong troubleshooting, analytical, communication, and collaboration skills.

Salary Range: 95,700 - 120,000

Benefits: Employees (and their eligible family members) may enroll in the following types of insurance coverage: medical, dental, vision, legal, and accidental death and dismemberment, as well as FSA/HSA (depending on enrolled medical plan). Yum! also provides short-term disability, long-term disability, and life insurance. Employees may enroll in our 401(k) plan. Yum! provides 4 weeks of vacation, paid sick leave, 10 paid holidays, a floating day off, half day Fridays year-round and 2 paid days for volunteer time each calendar year. To learn more about working at Yum! -

At Yum!, one of our core values is to Believe in ALL People. This means seeing the value in everyone and unlocking their full potential to be their best self. YUM! Brands, Inc. (including its subsidiaries Yum Restaurant Services Group, LLC (“YRSG”) and Yum Connect, LLC (“Yum Digital and Technology”)(collectively, “Yum”) is proud to be an equal opportunity employer and is committed to equity, inclusion, and belonging for all dimensions of diversity. We do not discriminate based on race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, disability status, age, or any other protected characteristic. Yum! is committed to working with and providing reasonable accommodation to applicants with disabilities or special needs.

US Job Seekers/Employees - to view the “Know Your Rights” poster and supplement and the Pay Transparency Policy Statement.

Stay connected with Yum! Brands and be the first to know about new opportunities across KFC, Taco Bell, The Habit Burger Grill, and our corporate functions.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer I
Site Reliability Engineer I

KFC • Plano (TX)

On-site
USD 96,000 - 120,000
Medical insurance
Dental insurance
Vision insurance
+3
Sr. Security Engineer
Sr. Security Engineer

Yum! Brands • Kansas

On-site
USD 117,000 - 148,000
Medical, dental, vision insurance
401(k) plan
Paid time off
+2
Sr. Hardware Integration Engineer
Sr. Hardware Integration Engineer

Yum! Brands • Kansas

On-site
USD 118,000 - 148,000
Medical, dental, vision insurance
401(k) plan
Paid vacation
+1
Automation Engineer
Automation Engineer

Yum! Brands • Plano (TX)

On-site
USD 99,000 - 117,000
Medical insurance
401(k) plan
Paid time off
Associate Manager, Learning & Development
Associate Manager, Learning & Development

YUM • Irvine (CA)

Hybrid
USD 109,000 - 129,000
Medical, dental, vision insurance
401(k) plan
4 weeks vacation + sick leave
+3
Sr. Security Engineer
Sr. Security Engineer

KFC Corporation • United States

On-site
USD 117,000 - 148,000
Insurance coverage: medical, dental, &
401(k) plan, vacation, holidays, sick/
Paid time off and volunteer days
Analyst, Treasury
Analyst, Treasury

YUM • Louisville (KY)

On-site
USD 89,000 - 100,000
4 weeks vacation
Paid holidays
Sick leave
+1
Associate Manager, Learning and Development
Associate Manager, Learning and Development

Yum! Brands • Irvine (CA)

On-site
USD 109,000 - 129,000
Medical, dental, vision insurance
401(k) plan
4 weeks vacation plus paid holidays
+1
Net Revenue Management Business Analyst
Net Revenue Management Business Analyst

KFC Corporation • Plano (TX)

On-site
USD 90,000 - 105,000
4 weeks vacation
Paid holidays
Volunteer days
+1
Sr. Technical Program Manager
Sr. Technical Program Manager

Yum! Brands • Plano (TX)

On-site
USD 118,600 - 139,400
Medical, dental, vision insurance
401(k) plan
Four weeks of vacation
+2