Position Overview
At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are united in delivering the best experience for our customers and fostering an inclusive workplace culture where all employees feel respected, valued, and have an opportunity to contribute to the company’s success. As a Software Engineering Manager for PNC's Site Reliability Engineering Center (SRC), you will work within the Information Technology Group and manage the daylight shift at one of our IT Hubs: Cleveland, Ohio; Birmingham, Alabama; Pittsburgh, Pennsylvania; Dallas, Texas; Denver, Colorado or Phoenix, Arizona. The SRC focuses on establishing a culture of operational excellence by ensuring infrastructure, platforms, and applications meet onboarding standards that improve reliability, enable proactive issue resolution, and reduce customer impact. This role supports the vision of building a collaborative technology organization across application, infrastructure, and security teams to deliver a stable, reliable, and secure environment.
Key Responsibilities
- Manage SRE and related teams; lead, coach, develop SRE engineers; set clear goals, drive accountability, and foster a culture of ownership and excellence.
- Lead incident management and remediation; manage end‑to‑end incident response for major (P1/P2) incidents, guide real‑time triage, diagnostics, and troubleshooting, and communicate with stakeholders.
- Provide technical leadership in production support; serve as an escalation point for complex issues across applications, infrastructure (Linux/Windows), databases (Oracle, SQL), middleware and integrations.
- Drive problem management and root‑cause resolution; lead root‑cause analysis (RCA) for recurring incidents, ensure ownership and resolution of problem records, and promote knowledge sharing via runbooks and knowledge articles.
- Oversee change management and release execution; validate change readiness, testing, rollback strategies, and risk assessments; represent the team in CAB reviews.
- Advance monitoring, alerting, and observability; lead efforts to build and optimize monitoring dashboards and alerting frameworks, champion tools such as Dynatrace, BigPanda, Logscale, and improve signal‑to‑noise ratio through tuning.
- Champion resiliency, stability, and availability; ensure high availability of critical systems, oversee disaster recovery, failover, and continuity testing, and drive MTTR improvement.
- Enable scalability and performance optimization; guide capacity planning and performance tuning strategies and partner with development teams for performance‑driven design improvements.
- Lead a 24x7 production support model; manage team participation in a 24x7 on‑call rotation, oversee incident bridges, war rooms, and escalation processes.
- Drive automation and operational efficiency; identify opportunities to reduce manual effort, implement automation across incident remediation, monitoring, deployment, and validation, and standardize runbooks.
- Ensure governance, risk, and compliance; maintain adherence to enterprise policies and regulatory standards, support audits, vulnerability remediation, and promote security and data governance practices.
Qualifications
- 5+ years of related experience and 3+ years of management experience.
- Strong experience in Site Reliability Engineering, Production Support, or DevOps.
- Proven ability to lead teams in high‑availability, enterprise environments.
- Deep understanding of incident, problem, and change management frameworks.
- Hands‑on knowledge of monitoring tools, cloud/infrastructure platforms, and automation.
- Experience improving system reliability, observability, and operational maturity.
- Strong communication skills and ability to lead during high‑pressure situations.
- Experience with OCP under infrastructure (Linux/Windows, OCP), MongoDB, Cassandra, Oracle, SQL, Elasticsearch, Redis, MQ and Kafka is a plus.
- Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent combination of education, certification, and experience.
- Preferred skills include Agile development, application delivery, coaching, and runbook development.
Benefits & Compensation
- Base salary range: $100,100.00 – $204,490.00 (may vary by location and experience).
- Incentive eligible rewards program.
- Medical, prescription drug, dental, and vision coverage; Health Savings Account.
- Life insurance, short‑term and long‑term disability protection.
- 401(k) with company match, pension and stock purchase plans.
- Dependent care reimbursement, back‑up child/elder care, adoption and surrogacy reimbursement.
- Educational assistance for eligible programs.
- Robust wellness program with financial incentives.
- Paid time off: maternity and/or parental leave; up to 11 paid holidays; 9 occasional absence days; 15‑25 vacation days depending on career level.
Equal Employment Opportunity (EEO)
PNC provides equal employment opportunity to qualified persons regardless of race, color, sex, religion, national origin, age, sexual orientation, gender identity, disability, veteran status, or other categories protected by law. This position is subject to the requirements of Section 19 of the Federal Deposit Insurance Act (FDIA) and, for any registered role, the Secure and Fair Enforcement for Mortgage Licensing Act of 2008 (SAFE Act) and/or the Financial Industry Regulatory Authority (FINRA).
Disability Accommodations Statement
If an accommodation is required to participate in the application process, please email AccommodationRequest@pnc.com with the subject line "accommodation request" and include your name, job ID, and preferred method of contact. Non‑related emails will not receive a response. All information provided will be kept confidential and used only to provide needed reasonable accommodations. PNC fosters an inclusive and accessible workplace and provides reasonable accommodations to qualified applicants and employees with disabilities.