Senior Director, Head, SRE and Production Operations

RBC

Toronto

On-site

CAD 180,000 - 240,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Bonuses
Stock options
Flexible benefits

Job summary

RBC is seeking a Head of SRE and Production Operations to lead the vision, design, and support of Site Reliability Engineering solutions across Commercial & Payments Technology within Technology & Operations. You will drive automated observability, end-to-end reliability, and cost-efficient operations.

The role demands 8–12 years of senior SRE/infra leadership, strong cloud, containerization, and AI-driven operations expertise, plus proven vendor and cross-functional leadership capabilities.

Qualifications

  • Minimum 8–12 years of experience leading SRE, production operations, or infrastructure engineering teams at scale.
  • Deep technical expertise in observability, monitoring, and incident management platforms.
  • Demonstrated experience with cloud platforms, containerization, and automation technologies.
  • Strong understanding of reliability engineering principles, SLOs, SLIs, and error budgets.
  • Proven track record managing vendor relationships and complex third-party technology partnerships.
  • Experience leading cross-functional teams and working across multiple business units.
  • Strong communication and relationship management abilities.
  • Demonstrated success implementing industry best practices and driving technical transformation.
  • Experience with AI and emerging technologies in operations and reliability engineering.

Responsibilities

  • Set the vision for SRE product offerings including monitoring, alerting, ML anomaly detection, self-healing, and reliability testing.
  • Drive technical solutions by tracking industry practices and applying them to RBC environment.
  • Lead thought leadership to align SRE tools with strategic objectives.
  • Run practice forums to review SRE solutions and align with enterprise vision.
  • Enable automated end-to-end observability across CPT applications.
  • Develop governance and cross-enterprise mindset for SRE.
  • Provide expertise, direction, coaching, and development to build team capability.
  • Attract, hire, and retain top talent; manage vendor relationships.

Skills

Observability
Monitoring
Incident management
Cloud platforms
Containerization
Automation
SRE principles (SLOs/SLIs)
Vendor management
Cross-functional leadership
Communication
AI in operations

Job description

What is the Opportunity? As Head of SRE and Production Operations, you will lead the vision, design, development, implementation, and support of Site Reliability Engineering (SRE) solutions for all applications supported by Commercial & Payments Technology (CPT) within Technology & Operations (T&O). This strategic role is instrumental in establishing world-class operational excellence, enabling automated end-to-end observability, and driving continuous reliability improvements across critical enterprise applications.

Job Description

What Will You Do?

Technical Leadership
  • Set the vision for SRE product offerings including monitoring, alerting, machine learning anomaly detection, self-healing capabilities, and reliability testing
  • Drive best-in-class technical solutions by tracking industry-leading practices and applying them to the RBC environment
  • Lead thought leadership and out-of-the-box thinking to ensure SRE tools and processes align with strategic objectives
  • Run engineering practice forums to facilitate the review of SRE solutions and maintain alignment with team and enterprise vision
  • Enable automated end-to-end observability across all critical applications within Commercial & Payments Technology (CPT)
  • Leverage unit, department, and enterprise-wide teams to develop better solutions and achieve a cross-enterprise mindset
Production Support
  • Perform production support role, including off-hours support responsibilities
  • Assist in incident management and problem management for applications in scope
  • Continuously evaluate what went well, what went wrong, and what can be done to improve and prevent issues in the future
  • Maintain technology currency through server patching, certificate renewal, and other maintenance activities with a keen eye on automating opportunities
  • Ensure availability and uptime of applications in scope according to defined service level objectives
  • Ensure compliance of all systems and applications in scope, including maintaining segregation of duties
  • Leverage AI and other key technologies to drive efficiencies and meet Commercial & Payments Technology (CPT), Technology & Operations (T&O) and respective business goals and KPIs
Strategy
  • Drive the overall SRE strategy within Innovation, Intelligent Operations, and partner groups, owning roadmap development
  • Lead the team through execution of the SRE roadmap for Innovation and partner groups
  • Drive participation in strategic steering groups
  • Lead the adoption of new technologies within SRE
  • Enable teams to meet objectives by implementing changes to processes, tools, and methods that result in increased agility and effective cost management
  • Develop governance of the unit in collaboration with executive leadership to provide structure and definition for effectively managing the Innovation SRE Organization
  • Drive industry best standards and practices for the most effective and efficient SRE and production support services
People Leadership
  • Provide expertise, direction, coaching, and development to build team capability and succession planning
  • Select and build a high-performing, diverse team that leverages individual capabilities and strengths
  • Promote a mindset for sustained success, growth, diversity, and an overall engineering mindset
  • Ensure that employees understand RBC’s vision and support and reinforce targeted behaviours that contribute to RBC goals
  • Provide focus and clarity in establishing individual goals, driving performance enablement, supporting career development, and rewarding strong performance
  • Attract, hire, and retain top talent
  • Manage key vendor relationships, contractors, and respective performance management to enable successful execution of SRE and production support services
What do you need to succeed?
Must-have
  • Minimum 8–12 years of experience leading SRE, production operations, or infrastructure engineering teams at scale
  • Deep technical expertise in observability, monitoring, and incident management platforms
  • Demonstrated experience with cloud platforms, containerization, and automation technologies
  • Strong understanding of reliability engineering principles, SLOs, SLIs, and error budgets
  • Proven track record managing vendor relationships and complex third-party technology partnerships
  • Experience leading cross-functional teams and working across multiple business units
  • Strong communication and relationship management abilities
  • Demonstrated success implementing industry best practices and driving technical transformation
  • Experience with AI and emerging technologies in operations and reliability engineering
Nice to Have
  • Payments industry experience
What’s in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
  • Leaders who support your development through coaching and managing opportunities
  • Ability to make a difference and lasting impact
  • Work in a dynamic, collaborative, progressive, and high-performing team
  • A world-class training program in financial services
  • Opportunities to do challenging work
  • Opportunities to take on progressively greater accountabilities
  • Opportunities to build close relationships with clients and stakeholders
Job Skills

Application Development, Application Maintenance, Applications Architecture, Commercial Acumen, Enterprise Application Delivery, Information Technology Management, Information Technology Trends, Programming Languages, System Applications

Additional Job Details

Address: RBC WATERPARK PLACE, 88 QUEENS QUAY W:TORONTO

City: Toronto

Country: Canada

Work hours/week: 37.5

Employment Type: Full time

Platform: TECHNOLOGY AND OPERATIONS

Job Type: Regular

Pay Type: Salaried

Posted Date: 2026-08-11

Application Deadline: 2026-09-14

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - SRE
Site Reliability Engineer - SRE

RBC • Toronto

On-site
CAD 110,000 - 140,000
Total rewards
Coaching & development
Impact
+3
Principal Engineer, Cyber Technology Operations SRE (Global Security)
Principal Engineer, Cyber Technology Operations SRE (Global Security)

RBC • Toronto

On-site
CAD 140,000 - 210,000
Total rewards program
Bonuses and stock where applicable
Flexible benefits
Director, Data Engineer
Director, Data Engineer

ODAIA • Toronto

On-site
CAD 150,000 - 230,000
Bonuses
Stock options
Flexible benefits
Senior Technical Systems Analyst
Senior Technical Systems Analyst

RBC • Toronto

On-site
CAD 90,000 - 120,000
Total rewards program
Bonuses & flexible benefits
Stock options where applicable
Senior Full Stack Engineer at RBC
Senior Full Stack Engineer at RBC

RBC • Toronto

On-site
CAD 110,000 - 150,000
Bonuses
Flexible benefits
Stock options where applicable
+1
Director, Data Engineer
Director, Data Engineer

RBC • Toronto

On-site
CAD 180,000 - 230,000
Bonus and incentive programs
Competitive compensation and stock
Comprehensive benefits
Director - Datacenter Network & Security Operations
Director - Datacenter Network & Security Operations

RBC • Toronto

On-site
CAD 180,000 - 260,000
Senior Application Support Analyst
Senior Application Support Analyst

RBC • Toronto

On-site
CAD 90,000 - 120,000
Lead Business Systems Analyst
Lead Business Systems Analyst

RBC • Toronto

On-site
CAD 90,000 - 120,000
Bonuses
Stock options
Flexible benefits
+1
Lead Platform Engineer
Lead Platform Engineer

RBC • Toronto

On-site
CAD 150,000 - 190,000
Total rewards
Flexible work hours
Training program
+1