CaaS Private Site Reliability Lead Engineer - Vice President

Deutsche Bank

Cary (NC)

Hybrid

USD 125,000 - 185,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Hybrid working model
Vacation, personal and volunteer days
Employee Resource Groups
Health and wellbeing benefits
Retirement savings plans
Parental leave and family building
Educational resources, matching gifts

Job summary

Deutsche Bank in Cary, NC seeks a CaaS Private Site Reliability Engineer (VP) to lead reliability, resilience, and operational excellence for the CaaS Private platform in the U.S. You will apply production engineering discipline to Kubernetes, observability, automation, and incident management while aligning with enterprise reliability standards.

You will partner with engineering, operations, and application teams to improve service health, strengthen platform readiness, and drive measurable

Qualifications

  • Extensive experience with Kubernetes, Linux, and distributed systems reliability.
  • Strong hands-on experience with observability, monitoring, alerting, and incident response.
  • Proven ability to design automation, self-healing workflows, and runbook improvements.

Responsibilities

  • Lead the reliability strategy for the CaaS Private platform in the US, including SLO frameworks and incident management maturity.
  • Drive resilience improvements across observability, capacity planning, upgrade safety, disaster readiness, and automation.
  • Own complex production issues, identify root causes, and implement preventive fixes.
  • Define service indicators, alert thresholds, dashboards, and production readiness criteria.

Skills

Kubernetes
Linux
Distributed systems reliability
Production platform operations
Observability
Incident management
SLOs and dashboards

Job description

Job Title

CaaS Private Site Reliability Engineer

Corporate Title

Vice President

Location

Cary, NC

Who we are

In short - an essential part of Deutsche Bank's technology solution, developing applications for key business areas.

Our Technologists drive Cloud, Cyber and business technology strategy while transforming it within a robust, hands‑on engineering culture. Learning is a key element of our people strategy, and we have a variety of options for you to develop professionally. Our approach to the future of work champions flexibility and is rooted in the understanding that there have been dramatic shifts in the ways we work.

Having first established a presence in the Americas in the 19th century, Deutsche Bank opened its US technology center in Cary, North Carolina in 2009. Learn more about us here.

Overview

As a CaaS Private Site Reliability Engineer, you will lead reliability, resilience, and operational excellence for the CaaS Private platform in the U.S. You will bring production engineering discipline to Kubernetes, observability, automation, and incident management while helping teams improve service health and platform readiness. You will partner across engineering, operations, and application teams to strengthen SLOs, reduce manual intervention, and ensure platform changes are measurable, supportable, and aligned to enterprise reliability standards.

What We Offer You
  • A diverse and inclusive environment that embraces change, innovation, and collaboration
  • A hybrid working model, allowing for in‑office / work from home flexibility, generous vacation, personal and volunteer days
  • Employee Resource Groups support an inclusive workplace for everyone and promote community engagement
  • Competitive compensation packages including health and wellbeing benefits, retirement savings plans, parental leave, and family building benefits
  • Educational resources, matching gift and volunteer programs
What You'll Do
  • Lead the reliability strategy for the CaaS Private platform in US, including SLO frameworks, operational standards, and incident management maturity
  • Drive resilience improvements across observability, capacity planning, upgrade safety, disaster readiness, and operational automation
  • Own complex production management issues by leading troubleshooting, identifying root causes, and implementing preventive fixes
  • Define and refine service indicators, alert thresholds, dashboard standards, production readiness criteria, escalation paths, and postmortem follow‑through
  • Develop automation and self‑healing workflows that reduce manual intervention, improve recovery times, and strengthen platform supportability
  • Partner with cross‑functional teams to ensure platform changes are measurable, supportable, and aligned with reliability objectives
How You'll Lead
  • Mentor junior and middle engineers while fostering a culture of blameless learning, measurable reliability, and operational excellence
  • Influence platform architecture and roadmap decisions using data from incidents, capacity models, operational trends, and reliability metrics
  • Communicate strategic insights and practical recommendations to engineering, operations, and business stakeholders to guide reliability investments
Skills You'll Need
  • Extensive experience with Kubernetes, Linux, distributed systems reliability, and production platform operations
  • Strong hands‑on experience with observability, monitoring, alerting, dashboarding, incident response, and root cause analysis
  • Proven ability to design and implement automation, self‑healing workflows, operational checks, maintenance tasks, and runbook improvements
  • Experience defining SLOs, service indicators, alert quality standards, production readiness practices, and escalation models
  • Ability to lead complex reliability improvements independently while partnering across engineering, operations, and application teams
Skills That Will Help You Excel
  • Strong communication skills with the ability to explain technical findings to technical and non‑technical stakeholders
  • Sound operational judgment with the ability to balance urgency, risk, and long‑term platform stability during incidents
  • Growth mindset with a focus on continuous learning, process improvement, and measurable reliability outcomes
  • Experience mentoring engineers and improving operational culture across platform or infrastructure teams
  • Background with automation tools, infrastructure‑as‑code practices, capacity planning, disaster readiness, or cloud‑native platform operations
Expectations

It is the Bank's expectation that employees hired into this role will work in the Cary, NC office in accordance with the Bank's hybrid working model.

Deutsche Bank provides reasonable accommodations to candidates and employees with a substantiated need based on disability and/or religion.

Salary

The salary range for this position in Cary is $125,000 to $185,000. Actual salaries may be based on a number of factors including, but not limited to, a candidate's skill set, experience, education, work location and other qualifications. Posted salary ranges do not include incentive compensation or any other type of remuneration.

Equal Opportunity and Legal Notices

We welcome applications from all people and promote a positive, fair and inclusive work environment.

Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status or other characteristics protected by law. The following notices apply: EEOC Know Your Rights; Employee Rights and Responsibilities under the Family and Medical Leave Act; and Employee Polygraph Protection Act.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Lead Engineer – Vice President (Kubernetes)
Site Reliability Lead Engineer – Vice President (Kubernetes)

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Generous vacation
Parental leave
Site Reliability Lead Engineer - Vice President (Kubernetes)
Site Reliability Lead Engineer - Vice President (Kubernetes)

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Generous vacation
Volunteer days
+3
CaaS Private Site Reliability Engineer - Assistant Vice President
CaaS Private Site Reliability Engineer - Assistant Vice President

Deutsche Bank • Cary (NC)

Hybrid
USD 100,000 - 153,000
Diverse and inclusive environment
Hybrid work model
Employee Resource Groups
+2
Platform Engineering & Site Reliability Engineering- Vice President
Platform Engineering & Site Reliability Engineering- Vice President

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid work model
Health benefits
Retirement savings plans
Senior SRE Lead & VP — Hybrid Platform Reliability
Senior SRE Lead & VP — Hybrid Platform Reliability

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Generous vacation
Parental leave
VP, CaaS Private Platform Reliability
VP, CaaS Private Platform Reliability

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Vacation, personal and volunteer days
Employee Resource Groups
+4
Kubernetes Platform SRE Lead
Kubernetes Platform SRE Lead

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Generous vacation
Volunteer days
+3
Senior Engineer - Vice President
Senior Engineer - Vice President

Deutsche Bank • Cary (NC)

Hybrid
USD 125,000 - 185,000
Hybrid working model
Vacation and personal days
Employee Resource Groups
+4
Technical Project / Program Manager - Network Security & Cyber Resilience - Assistant Vice President
Technical Project / Program Manager - Network Security & Cyber Resilience - Assistant Vice President

Deutsche Bank • Cary (NC)

Hybrid
USD 100,000 - 153,000
Hybrid work model
Vacation & personal days
Health benefits
+2
Senior Production Support Engineer - Assistant Vice President
Senior Production Support Engineer - Assistant Vice President

Deutsche Bank • Cary (NC)

Hybrid
USD 100,000 - 153,000
Hybrid work model
Flexible vacation & personal days
ERGs and community engagement
+1