Senior Site Reliability Engineer

Comcast

Centennial (CO)

On-site

USD 104,000 - 157,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical & Dental
401(k) Savings Plan
Generous paid time off
Life Milestones – adoption, childcare,
Courtesy Services – free digital TV &

Job summary

Comcast Technology Solutions, a software technology company headquartered in Denver, Colorado, USA, seeks an experienced Reliability Engineer to ensure scalable, robust services across cloud platforms. You will monitor infrastructure, automate workflows, and participate in incident response, with a focus on reducing toil and improving observability.

The role emphasizes problem solving, collaboration with cross-functional teams, and delivering high-quality solutions for streaming, content, and

Qualifications

  • Understanding of wider operational performance factors influenced by infrastructure workload.
  • Strong investigative mindset and attention to detail.
  • Ability to diagnose problems and implement permanent fixes.
  • Automation as a means to address scale challenges.

Responsibilities

  • Analyze and forecast system capacity for scalability and performance.
  • Participate in incident response and post-incident reviews.
  • Develop and maintain monitors and alerts across services.
  • Tune and configure to optimize system performance.
  • Develop disaster recovery plans to ensure business continuity.
  • Create documentation for systems and processes.
  • Monitor cloud infrastructure for efficient resource use.
  • Identify opportunities for automation and implement solutions.
  • Collaborate with cross-functional teams to deliver high-quality solutions.
  • May direct workflow as a technical lead.
  • Exercise independent judgment in critical matters.

Skills

Cloud platforms
Scripting
IaC
Docker/Kubernetes
Monitoring
DB performance
Kubernetes monitoring
Communication
Problem solving

Education

Bachelor’s Degree

Tools

Docker
Kubernetes
Terraform
Ansible
Datadog
Splunk
NoSQL/SQL

Job description

Make your mark at Comcast — a Fortune 30 global media and technology company. From the connectivity and platforms we provide, to the content and experiences we create, we reach hundreds of millions of customers, viewers, and guests worldwide. Become part of our award-winning technology team that turns big ideas into cutting-edge products, platforms, and solutions that our customers love. We create space to innovate, and we recognize, reward, and invest in your ideas, while ensuring you can proudly bring your authentic self to the workplace. Join us. You’ll do the best work of your career right here at Comcast. (In most cases, Comcast prefers to have employees on-site collaborating unless the team has been designated as virtual due to the nature of their work. If a position is listed with both office locations and virtual offerings, Comcast may be willing to consider candidates who live greater than 100 miles from the office for the remote option.)

Job Summary

COMCAST Technology Solutions is a software technology company headquartered in Denver, Colorado, USA. We enable streaming services, TV stations, pay TV operators, content providers, broadband media sites, and mobile businesses to solve their unique media management and video publishing requirements. Our Cloud Video Platform (CVP), provided as a service, offers a diverse product catalogue. By leveraging Comcast CVP, our customers can securely manage their digital media, publish content to various IP devices, and effectively monetize their distribution directly to consumers. Our proven media management and publishing technology provides a versatile approach to meet each customer’s unique business requirements and scales fluidly to support their growth. Our customers include Deutsche Telekom, Viaplay, Fox, Disney, NBC, Paramount , and many others. Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless and robust services to our users. We’re a distributed team of engineers with diverse skillsets who thrive on solving complex challenges and driving innovation with a focus on improving observability and reducing toil.

Job Description
Core Responsibilities
  • Analyzes and forecasts system capacity requirements to ensure scalability and performance for high-profile events.
  • Participate in incident response efforts, conduct post-incident reviews, and implement corrective actions and improvements to monitoring.
  • Develop and maintain monitors and alerts across all services.
  • Optimize system performance through tuning and configuration adjustments.
  • Develop and maintain disaster recovery plans and procedures to ensure business continuity.
  • Create and maintain comprehensive documentation for systems, processes, and procedures.
  • Monitor and optimize cloud infrastructure to ensure efficient resource utilization.
  • Identify opportunities for automation and implement solutions to reduce manual intervention.
  • Work closely with cross-functional teams to align on goals and deliver high-quality solutions.
  • Does not have any direct supervisory responsibilities. May direct workflow and act as a technical lead.
  • Consistently exercises independent judgment and discretion in matters of significance.
  • Shows regular, consistent and punctual attendance.
  • Other duties and responsibilities as assigned.
ABOUT YOU

Our people are the most important part of our business. We are fundamentally looking for forward-thinking, enthusiastic problem solvers. People who love a challenge, constantly evaluate and question, and, above all, love to ship a product that solves real problems. While these characteristics outweigh any specific technical skills, you should be able to demonstrate some of the following

  • An understanding of wider operational performance factors influenced by the underlying infrastructure workload, such as server platforms, databases and networking.
  • A strong drive to be a ‘detective’ and understand why things are working (or not working) as they should, in other words, a passion for detail and an investigative nature.
  • The ability to proactively diagnose problems using your holistic knowledge-set — and then get busy with coding a permanent fix, rewriting a process or working with third parties to ensure that lessons are learned, and problems never recur.
  • A vision of automation as an opportunity to overcome scale challenges, and a flexible approach to technologies.
Must Have Skills
  • Strong knowledge of cloud platforms (e.g., AWS, GCP, Azure).
  • Proficiency in scripting languages (e.g., Python, Bash).
  • Experience with infrastructure-as-code tools (e.g., Terraform, Ansible).
  • Familiarity with containerization and orchestration tools (e.g., Docker, Kubernetes).
  • Experience with monitoring tools (e.g., Datadog, Splunk)
  • Experience with Database performance monitoring and tuning (e.g., NoSQL, SQL)
  • Experience with Kubernetes performance monitoring and tuning
  • Excellent problem-solving skills and attention to detail.
  • Effective communication and collaboration skills.
Desirable Skills
  • CI/CD pipeline management
  • Cloud Cost Optimization
  • Automating deployment processes
  • Front-end development (e.g., React)
Here’s a look at just some of the perks and benefits we make available to our US-based employees:
  • Medical & Dental
  • 401(k) Savings Plan
  • Generous paid time off
  • Life Milestones – from adoption assistance, childcare resources, pet insurance, and more, Comcast supports you at all life stages.
  • Courtesy Services – We offer all of our full-time employees in serviceable areas free digital TV and internet.

Learn more at jobs.comcast.com/life-at-comcast/benefits

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive other benefits and privileges of employment. Please contact us to request an accommodation.

  • This information has been designed to indicate the general nature and level of work performed by employees in this role. It is not designed to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and qualifications.
  • Comcast is an EOE/Veterans/Disabled/LGBT employer.
  • Discount tickets for Universal Resorts, including theme park tickets and onsite hotel rooms.

Comcast is an equal opportunity workplace. We will consider all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender identity, national origin, disability, veteran status, genetic information, or any other basis protected by applicable law.

Skills

Reliability Engineering; Site Reliability Engineering; Reliability Analysis

Salary

Primary Location Pay Range: $104,431.52 – $156,647.28

Comcast intends to offer the selected candidate base pay within this range, dependent on job-related, non-discriminatory factors such as experience. The application window is 30 days from the date job is posted, unless the number of applicants requires it to close sooner or later.

The application window is 30 days from the date job is posted, unless the number of applicants requires it to close sooner or later.

Base pay is one part of the Total Rewards that Comcast provides to compensate and recognize employees for their work. Most sales positions are eligible for a Commission under the terms of an applicable plan, while most non-sales positions are eligible for a Bonus. Additionally, Comcast provides best-in-class Benefits to eligible employees. We believe that benefits should connect you to the support you need when it matters most, and should help you care for those who matter most. That’s why we provide an array of options, expert guidance and always-on tools, that are personalized to meet the needs of your reality – to help support you physically, financially and emotionally through the big milestones and in your everyday life. Please visit the compensation and benefits summary (https://jobs.comcast.com/benefits) on our careers site for more details.

Education

Bachelor’s Degree

While possessing the stated degree is preferred, Comcast also may consider applicants who hold some combination of coursework and experience, or who have extensive related professional experience.

Relevant Work Experience

5-7 Years

Job Family Group

Engineering

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Blueface Ltd • Centennial (CO)

Hybrid
USD 104,000 - 157,000
Medical & Dental 401(k)
Generous paid time off
Life Milestones
+2
Software Engineer 3 - Kubernetes Platform Management - Freewheel
Software Engineer 3 - Kubernetes Platform Management - Freewheel

Comcast • Reston (VA)

On-site
USD 120,000 - 180,000
Principal Software Development Engineer
Principal Software Development Engineer

Comcast • Pennsylvania

On-site
USD 120,000 - 180,000
Medical & Dental
401(k) Savings Plan
Paid time off
+3
Data Center Operations and Design Engineer
Data Center Operations and Design Engineer

Comcast • Hillsboro (OR)

On-site
USD 90,000 - 130,000
Medical & Dental
401(k) Savings Plan
Generous paid time off
+4
X-Labs Design Engineer
X-Labs Design Engineer

Comcast • Philadelphia

On-site
USD 114,000
Comprehensive benefits package
Career growth and development opportunities
Flexible work schedules
Sr. Software Engineer - Reston, Hybrid - FreeWheel
Sr. Software Engineer - Reston, Hybrid - FreeWheel

Comcast • Reston (VA)

On-site
USD 142,000 - 213,000
Critical Infrastructure Engineer
Critical Infrastructure Engineer

Blueface Ltd • Centennial (CO)

On-site
USD 77,000 - 117,000
CTS Solutions and Integrations Engineer
CTS Solutions and Integrations Engineer

Cloudjobs • West Virginia

On-site
USD 115,000 - 172,000
Medical & Dental
401k Savings Plan
Generous paid time off
+2
Critical Infrastructure Engineer
Critical Infrastructure Engineer

Comcast • Centennial (CO)

On-site
USD 77,936 - 116,904
Sr. Manager, Data Center Engineering
Sr. Manager, Data Center Engineering

Blueface Ltd • Chester

On-site
USD 140,000 - 190,000