Senior Site Reliability Engineer

RBC

Minneapolis (MN)

On-site

USD 80,000 - 140,000

Full time

35 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its SRE Team in Minneapolis. The role emphasizes reliability, automation, and AI-enhanced operations across hybrid cloud platforms.

You will design scalable SRE solutions, implement modern observability, and drive incident response improvements while collaborating with development, infrastructure, and support teams to maintain high availability for critical wealth management applications.

Qualifications

  • 5+ years in SRE, Prod/DevOps or related roles with strong ops depth.
  • Bachelor’s degree in CS/Engineering or equivalent experience.
  • Hands-on automation and configuration management with Ansible.
  • Scripting in Bash, Python, PowerShell or similar.
  • Experience with observability tooling and incident management.
  • Knowledge of SLIs/SLOs and reliability practices.
  • Cloud-native concepts and AI/ML in observability context.

Responsibilities

  • Build and enhance the SRE product base with monitoring and automated remediation.
  • Implement modern observability across applications with dashboards and alerts.
  • Design ML-based anomaly detection and self-healing solutions.
  • Standardize telemetry and instrumentation across platforms.
  • Automate workflows using Ansible, GitHub Actions and scripting.
  • Define and track service health metrics and runbooks.
  • Collaborate with development teams to ensure reliability pre/post deployment.
  • Lead incident management and root cause analysis.
  • Drive continuous improvement with AI-driven ops.

Skills

SRE experience
Automation scripting
Observability tooling
Incident management
Cross-team collaboration

Education

Bachelor's degree in CS/Engineering

Tools

Elasticsearch
Dynatrace
GitHub Actions
Kubernetes
OpenShift
Kafka
Ansible
Moogsoft
PagerDuty

Job description

Job Description

RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and platforms that support the wealth management business. Working at the intersection of software engineering, cloud-native operations, observability, and automation, the team plays a central role in delivering reliable digital services for both internal users and clients.

As a Senior Site Reliability Engineer, you will bring an engineering‑first mindset, strong operational judgment, and a passion for automation to improve system reliability at scale. You will work closely with development, infrastructure, platform, and support teams to build modern observability practices, improve incident response, strengthen reliability engineering standards, and drive the evolution toward intelligent, self‑healing operations.

This role is ideal for a hands‑on engineer who is equally comfortable improving production resilience, building automation, defining service‑level objectives, and shaping the future of AI‑enhanced operations. You will help design and implement scalable SRE solutions across the technology estate using tools and platforms such as Elasticsearch, Ansible, GitHub Actions, Dynatrace, PagerDuty, Moogsoft, Kubernetes, OpenShift, Kafka, and emerging AIOps capabilities.

What will you do?
  • Build and enhance the SRE product base – Develop intelligent monitoring, alerting, reliability testing, anomaly detection, and automated remediation capabilities.
  • Implement modern observability practices – Deploy metrics, logs, traces, dashboards, and actionable alerting across supported applications.
  • Design ML‑based anomaly detection and self‑healing solutions – Shift from reactive to predictive operations with automated issue remediation and appropriate governance controls.
  • Standardize telemetry and instrumentation – Improve visibility, coverage, and correlation of operational signals across platforms.
  • Automate operational workflows – Use Ansible, GitHub Actions, and scripting (Bash, Python, PowerShell) to streamline platform tasks and develop custom tooling.
  • Define and track service health metrics – Establish and improve SLIs, SLOs, error budgets, and evolve runbooks into automation‑first remediation patterns.
  • Partner with development teams – Ensure applications meet reliability and performance standards before and after deployment through close collaboration.
  • Lead incident and problem management – Troubleshoot production issues across all layers, participate in on‑call rotation, and drive root cause analysis and corrective actions.
  • Drive continuous improvement – Identify opportunities to simplify, automate, and modernize operations using engineering and AI‑driven approaches
What do you need to succeed?
Must-have
  • 5+ years of experience in Site Reliability Engineering, Production Engineering, DevOps, Platform Engineering, or Systems Engineering roles with strong operational depth.
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • Strong experience with infrastructure automation and configuration management, particularly Ansible.
  • Strong scripting and automation skills in Bash, Python, PowerShell, or similar languages.
  • Hands‑on experience with modern reliability and observability tooling such as Elasticsearch, Dynatrace, GitHub, Kubernetes, OpenShift, Kafka, PagerDuty, Moogsoft, or related platforms.
  • Strong understanding of production operations, incident management, root cause analysis, and reliability engineering practices.
  • Experience defining and operating SLIs, SLOs, alerting strategies, and service health metrics.
  • Knowledge of cloud‑native and distributed systems concepts, including resiliency, scalability, fault isolation, and performance tuning.
  • Understanding of AIOps, AI/ML concepts, or intelligent automation as applied to observability and operations.
  • Ability to work across teams, influence engineering practices, and communicate clearly with technical and non‑technical stakeholders.
Nice-to-have
  • Experience in financial services, wealth management, banking, insurance, or other highly regulated environments.
  • Experience with OpenTelemetry and telemetry standardization across distributed systems.
  • Hands‑on experience with Prometheus, Grafana, Splunk, Catchpoint, Azure Automation, or similar SRE and observability platforms.
  • Experience with CI/CD and developer platform tools such as Jenkins, Artifactory, and Vault.
  • Familiarity with containerization and cloud platform patterns, including Docker and Kubernetes‑based deployments.
  • Experience building or operating anomaly detection, predictive alerting, or self‑healing automation solutions.
  • Familiarity with AI governance, model validation, and operational controls in regulated environments.
What’s in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
  • Leaders who support your development through coaching and managing opportunities
  • Ability to make a difference and lasting impact
  • Work in a dynamic, collaborative, progressive, and high‑performing team
  • A world‑class training program in financial services

Expected salary range for this particular position $80,000-$140,000, depending on your experience, skills, and registration status, market conditions and business needs.

You have the potential to earn more through RBC’s discretionary variable compensation program which gives you an opportunity to increase your total compensation, provided the business meets its performance targets and you meet your individual goals.

RBC’s compensation philosophy and principles recognize the importance of a highly qualified global workforce and plays a critical role in attracting, engaging and retaining talent that:

  • Drives RBC’s high‑performance culture.
  • Enables collective achievement of our strategic goals.
  • Generates sustainable shareholder returns and above market shareholder value.
Job Skills

Agile Methodology, Application Infrastructure, Group Problem Solving, IT Automation, IT Monitoring, Operations Support, Production Support, Software Development Life Cycle (SDLC), Software Engineering, Software Product Technical Knowledge, System Applications, Systems Software

Additional Job Details

Address: 250 NICOLLET MALL:MINNEAPOLIS

City: Minneapolis

Country: United States of America

Work hours/week: 40

Employment Type: Full time

Platform: TECHNOLOGY AND OPERATIONS

Job Type: Regular

Pay Type: Salaried

Posted Date: 2026-07-13

Application Deadline: 2026-09-11

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Application Manager
Senior Application Manager

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 100,000 - 170,000
Total rewards program
Stock options where applicable
World-class training
Software Engineer
Software Engineer

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 80,000 - 140,000
Senior Application Manager
Senior Application Manager

RBC • Minneapolis (MN)

On-site
USD 100,000 - 170,000
Specialist, Operations I - Retirement Plan Operations
Specialist, Operations I - Retirement Plan Operations

RBC • Minneapolis (MN)

On-site
USD 50,000 - 80,000
Operations Specialist I - Trade Support
Operations Specialist I - Trade Support

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 50,000 - 80,000
Total Rewards Program
Flexible work-life balance
Leadership development opportunities
+2
Specialist, Operations I - Retirement Plan Operations
Specialist, Operations I - Retirement Plan Operations

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 50,000 - 80,000
Senior Payment Operations Manager
Senior Payment Operations Manager

RBC • Minneapolis (MN)

On-site
USD 90,000 - 160,000
Director, Operations Business Administration
Director, Operations Business Administration

RBC • Minneapolis (MN)

On-site
USD 120,000 - 180,000
Bonuses
Stock options
Flexible benefits
+1
Senior Payment Operations Manager
Senior Payment Operations Manager

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 90,000 - 160,000
Total Rewards
Flexible benefits
Coaching & development
+1
HR Data Systems Analyst
HR Data Systems Analyst

RBC Capital Markets, LLC • Minneapolis (MN)

On-site
USD 55,000 - 95,000
401(k) with company matching
Health, dental, vision, life, and long
Paid time off