Software Engineering Manager - Site Reliability Center

PNC

Pittsburgh (Allegheny County)

On-site

USD 115,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical and prescription drug coverage
401(k) with PNC match
Paid time off and holiday benefits
Educational assistance

Job summary

PNC is seeking a Software Engineering Manager for its Site Reliability Engineering Center. This position involves managing teams and leading incident management efforts, while ensuring system reliability and performance optimization.

The candidate will be responsible for overseeing SRE practices, fostering a culture of excellence, and providing technical guidance within IT hubs located across the United States, including Pittsburgh, Dallas, and more.

Qualifications

  • 5+ years of experience and 3+ years of management experience.
  • Strong experience in Site Reliability Engineering or DevOps.
  • Proven ability to lead teams in enterprise environments.

Responsibilities

  • Manage SRE teams; foster a culture of ownership and excellence.
  • Guide incident management for major incidents.
  • Provide technical leadership in production support.

Skills

Site Reliability Engineering
Production Support
DevOps
Strong Communication Skills
Incident Management
Problem Management
Change Management
Monitoring Tools
Linux/Windows

Education

5+ years of related experience
3+ years of management experience

Tools

Oracle
SQL
MongoDB
Cassandra
Elasticsearch
Redis
MQ
Kafka

Job description

Position Overview

At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. As a Software Engineering Manager for PNC's Site Reliability Engineering Center, you will work within PNC's Information Technology Group and be located at one of our IT Hubs: Cleveland, Ohio; Birmingham, Alabama; Pittsburgh, Pennsylvania; Dallas, Texas; Denver, Colorado or Phoenix, Arizona. You will manage the daylight shift.

Responsibilities
  • Manage SRE and related teams; lead, coach, and develop a team of SRE engineers; set clear goals, drive accountability, and foster a culture of ownership and excellence; partner with cross‑functional stakeholders to align technology and business objectives; support talent development, performance management, and succession planning; encourage innovation, continuous learning, and DevOps/SRE best practices.
  • Lead incident management & remediation; manage and actively participate in end‑to‑end incident response for major (P1/P2) incidents; guide real‑time triage, diagnostics, and troubleshooting across application, infrastructure, and network layers; ensure rapid execution of remediation actions and service restoration; provide clear, timely communication to stakeholders during incidents; oversee post‑incident analysis, reporting, and documentation to drive improvements.
  • Provide technical leadership in production support; serve as an escalation point for complex production issues; guide troubleshooting across applications, infrastructure (Linux/Windows), databases (Oracle, SQL), middleware, and integrations; ensure efficient log, metric, and system analysis; oversee batch/ETL monitoring and recovery processes; foster strong collaboration across engineering, infrastructure, and vendor teams.
  • Drive problem management & root‑cause resolution; lead root‑cause analysis efforts for major and recurring incidents; ensure ownership and resolution of problem records; drive permanent fixes and systemic improvements to eliminate repeat issues, identify trends and patterns to reduce risk and improve stability; partner with engineering teams to resolve code defects and system gaps and promote knowledge sharing via runbooks, knowledge articles, and error catalogs.
  • Oversee change management & release execution; ensure safe and compliant execution of production changes and releases; validate change readiness, testing, rollback strategies, and risk assessments; represent the team in CAB reviews, providing technical risk evaluation; oversee post‑implementation reviews (CPIR) and ensure follow‑through and drive improvements in change success rate and reduction in production defects.
  • Advance monitoring, alerting & observability; lead efforts to build and optimize monitoring, dashboards, and alerting frameworks, champion use of tools such as Dynatrace, BigPanda, Logscale, and enterprise platforms; improve signal‑to‑noise ratio through alert tuning; enable proactive issue detection before customer impact; strengthen event management and observability practices.
  • Champion resiliency, stability & availability; lead efforts to ensure high availability of critical systems; oversee disaster recovery, failover, and continuity testing; identify and eliminate single points of failure and drive improvements in MTTR, uptime, and service reliability.
  • Enable scalability & performance optimization; guide capacity planning and performance tuning strategies; ensure systems scale effectively under peak demand; partner with development teams for performance‑driven design improvements; optimize system configurations to improve efficiency and throughput.
  • Lead a 24x7 production support model; manage team participation in a 24x7 on‑call rotation; oversee engagement in incident bridges, war rooms, and escalations; support pod‑based operating models aligned to key applications; ensure seamless handoffs and global support continuity.
  • Drive automation & operational efficiency; identify and prioritize opportunities to reduce manual effort through automation; implement automation across incident remediation, monitoring and alerting, deployment and validation; promote standardized runbooks and automation frameworks and improve operational metrics and reduce toil.
  • Ensure governance, risk & compliance; maintain adherence to enterprise policies and regulatory standards; support audits, vulnerability remediation, and risk controls; ensure accurate documentation and operational procedures and champion security, access management, and data governance practices.
Qualifications
  • 5+ years of related experience and 3+ years of management experience.
  • Strong experience in Site Reliability Engineering, Production Support, or DevOps.
  • Proven ability to lead teams in high‑availability, enterprise environments.
  • Deep understanding of incident, problem, and change management frameworks.
  • Hands‑on knowledge of monitoring tools, cloud/infrastructure platforms, and automation.
  • Experience improving system reliability, observability, and operational maturity.
  • Strong communication skills with the ability to lead during high‑pressure situations.
  • Experience with Linux/Windows operating systems, OCP, Oracle, SQL, MongoDB, Cassandra, Elasticsearch, Redis, MQ, and Kafka is a plus.
Benefits
  • Medical and prescription drug coverage with Health Savings Account.
  • Dental and vision coverage.
  • Life insurance for employees and spouses/children.
  • Short‑ and long‑term disability protection.
  • 401(k) with PNC match.
  • Pension and stock purchase plans.
  • Dependent care reimbursement.
  • Backup child/elder care.
  • Adoption, surrogacy, and doula reimbursement.
  • Educational assistance including selected programs fully paid.
  • Robust wellness program with financial incentives.
  • Paid time off: maternity/parental leave, up to 11 holidays, 9 occasional absence days, 15–25 vacation days based on career level, and years of service accrual.
Equal Employment Opportunity

PNC provides equal employment opportunity to qualified persons regardless of race, color, sex, religion, national origin, age, sexual orientation, gender identity, disability, veteran status, or other categories protected by law. This position is subject to the requirements of Section 19 of the FDIA and, for any registered role, the SAFE Act and/or FINRA, which prohibit the hiring of individuals with certain criminal history.

Disability Accommodations Statement

If an accommodation is required to participate in the application process, please contact us via email at AccommodationRequest@pnc.com. Include “accommodation request” in the subject line title and be sure to provide your name, the job ID, and your preferred method of contact in the body of the email. Applicants may also call 877‑968‑7762 and say "Workday" for accommodation assistance. All information provided will be kept confidential and will be used only to the extent required to provide needed reasonable accommodations.
At PNC we foster an inclusive and accessible workplace. We provide reasonable accommodations to employment applicants and qualified individuals with a disability who need an accommodation to perform the essential functions of their positions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineering Manager - Site Reliability Center
Software Engineering Manager - Site Reliability Center

PNC Financial Services Group, Inc. • Pittsburgh

On-site
USD 100,100 - 204,490
Medical, prescription drug, dental, and vision coverage
401(k) with company match
Paid time off including maternity leave and vacation days
Software Engineering Manager-Site Reliability Center-Twilight
Software Engineering Manager-Site Reliability Center-Twilight

Fairygodboss • Birmingham (AL)

On-site
USD 101,000 - 205,000
401(k) matching
Health insurance
Paid time off
Site Rel Engineer Sr. - Model Integration Platform
Site Rel Engineer Sr. - Model Integration Platform

Fairygodboss • Denver (CO)

On-site
USD 86,000 - 158,000
Medical/prescription drug coverage
Dental and vision options
401(k) with company match
+2
System Reliability & Support Specialist Sr. - Site Reliability Center
System Reliability & Support Specialist Sr. - Site Reliability Center

PNC • Pittsburgh

On-site
USD 55,000 - 124,000
Medical insurance
Dental & Vision
401(k) with match
+1
Site Reliability Engineer
Site Reliability Engineer

Vibrant Pittsburgh • Pittsburgh

On-site
USD 90,000 - 120,000
Medical and prescription drug coverage
401(k) with PNC match
Paid time off including vacation and holidays
+1
Site Reliability Engineer — Automation & Resilience at Scale
Site Reliability Engineer — Automation & Resilience at Scale

Vibrant Pittsburgh • Pittsburgh

On-site
System Reliability & Support Specialist - Production Support
System Reliability & Support Specialist - Production Support

Fairygodboss • Phoenix (AZ)

On-site
USD 63,000 - 117,000
Medical and prescription drug coverage
Health Savings Account
Dental and vision options
+5
Site Reliability Engineer Sr.
Site Reliability Engineer Sr.

PNC • Cleveland (OH)

On-site
USD 86,000 - 144,000
Competitive salary
In-office role
Benefits package
+1
Software Engineer Lead - Site Reliability Engineering Center
Software Engineer Lead - Site Reliability Engineering Center

Fairygodboss • Lakewood (CO)

On-site
USD 86,000 - 159,000
Medical coverage
Dental and vision
401(k) with match
+2
Software Engineer Lead - Site Reliability Engineering Center
Software Engineer Lead - Site Reliability Engineering Center

Fairygodboss • Farmers Branch (TX)

On-site
USD 86,000 - 159,000
Health insurance
401(k) with PNC match
Pension Plan
+1