Cloud Engineer

RBC

Toronto

On-site

CAD 120,000 - 160,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

RBC in Toronto is seeking a Senior Engineer for Hybrid Infrastructure Services to design and operate scalable hybrid cloud platforms supporting internal application teams. You will work with modern tech like Kafka, Elasticsearch, Redis and leverage AI-assisted coding to accelerate infra development.

Responsibilities include building task-specific agents for on-call triage, onboarding validation, capacity forecasting, and ensuring security, monitoring, and compliance across environments.

Qualifications

  • Passion for debugging complex systems at scale.
  • Experience with AI-assisted coding tools in daily workflow.
  • Scripting and automation skills (Go, Shell, Ansible, Python) for infrastructure as code.
  • Experience with container technologies (Docker, Kubernetes).
  • Excellent communication and organizational skills.
  • Eagerness to learn and advance AI fluency.

Responsibilities

  • Collaborate with cross-functional teams to design and implement scalable, resilient hybrid Cloud services that align with organizational goals.
  • Use AI-assisted coding (Copilot, Claude Code, or equivalent) as your default way of writing and reviewing infrastructure code, scripts, and documentation — holding it to the same review bar as any other change, just produced faster.
  • Design, build, and operate task-specific agents that reduce manual toil — for example, on-call triage agents that correlate metrics and past incidents, onboarding agents that validate requests against policy, or capacity-forecasting agents that flag drift weeks ahead.
  • Define guardrails for agentic workflows: what an agent can act on autonomously versus what requires human sign-off, with full audit trails for every automated action.
  • Conduct performance analysis and optimization of services to ensure efficiency and cost-effectiveness.
  • Identify and implement best practices for resource utilization and scalability across diverse infrastructure landscapes.

Skills

AI-assisted engineering
Debugging at scale
Automation scripting
Go
Python
Shell

Tools

Docker
Kubernetes
Kafka
Elasticsearch
Redis
Copilot
Claude Code

Job description

What is the opportunity?

Join our dynamic and innovative Hybrid Infrastructure Services Team, where cutting-edge technology meets a passion for driving organizational change. We operate critical shared platforms including Kafka, Elasticsearch, and Redis supporting a large number of internal application teams.

Job Description
What is the opportunity?

Join our dynamic and innovative Hybrid Infrastructure Services Team, where cutting-edge technology meets a passion for driving organizational change. We operate critical shared platforms including Kafka, Elasticsearch, and Redis supporting a large number of internal application teams.

We're not looking for someone who just knows Kubernetes, Kafka, Elasticsearch, and Redis. We're looking for a Senior Engineer who treats AI-assisted engineering as a core skill — someone who uses tools like Copilot and Claude Code as a default way of working, and who can go further: designing and deploying task-specific agents that handle on-call triage, onboarding validation, upgrade pre-checks, and capacity forecasting at scale. You'll help shape how our team delivers meaningfully more value faster automation, smarter self-service, quicker incident response without the tradeoffs of understanding the systems less deeply.

What will you do?
Engineering
  • Collaborate with cross-functional teams to design and implement scalable, resilient hybrid Cloud services that align with organizational goals.
  • Use AI-assisted coding (Copilot, Claude Code, or equivalent) as your default way of writing and reviewing infrastructure code, scripts, and documentation — holding it to the same review bar as any other change, just produced faster.
  • Design, build, and operate task-specific agents that reduce manual toil — for example, on-call triage agents that correlate metrics and past incidents, onboarding agents that validate requests against policy, or capacity-forecasting agents that flag drift weeks ahead.
  • Define guardrails for agentic workflows: what an agent can act on autonomously versus what requires human sign-off, with full audit trails for every automated action.
  • Conduct performance analysis and optimization of services to ensure efficiency and cost-effectiveness.
  • Identify and implement best practices for resource utilization and scalability across diverse infrastructure landscapes.
Security and Compliance
  • Implement and maintain security measures to safeguard on-premise and public Cloud environments and data.
  • Ensure compliance with industry standards and regulatory requirements across deployment models, including explainability and auditability of any AI- or agent-driven action on production infrastructure.
Monitoring and Troubleshooting
  • Set up monitoring and alerting for proactive issue identification across on-premise and public Cloud environments.
  • Respond to incidents promptly, using AI-assisted correlation to speed up diagnosis — while owning the judgment calls (root cause, failover decisions, customer communication) that stay firmly human.
What do you need to succeed?
Must-have
  • Passion for debugging complex systems and an eye for problems that occur at scale.
  • Deep care for the resiliency of systems and the quality of what you ship — including AI-generated changes, which get the same scrutiny as hand‑written ones.
  • Hands‑on experience using AI coding assistants (Copilot, Claude Code, or similar) as part of your daily engineering workflow, not as a novelty.
  • Scripting and automation skills (e.g., Go, Shell, Ansible, Python) for infrastructure as code.
  • Experience with container technologies (e.g., Docker, Kubernetes).
  • Excellent communication and organizational skills.
  • Strong desire to learn new skills and technologies — this role expects continuous growth up a defined AI‑fluency ladder (from daily AI‑assisted engineering, to building personal workflow agents, to deploying production task‑specific agents).
Nice-to-have
  • Experience building or deploying LLM‑based agents or tools — e.g., tool‑calling/function‑calling patterns, retrieval‑augmented log/metric correlation, intent classification for request triage.
  • Software development experience in Java or Rust.
  • Familiarity with Elasticsearch, Kafka, or Redis, either from a development or operations perspective.
  • Understanding of AI governance, model risk, or LLM failure modes (hallucination, staleness, prompt injection) in a regulated environment.
Preferred certifications
  • Microsoft Certified: Azure Administrator Associate (or equivalent Azure/Cloud administration certification).
  • Certified Kubernetes Administrator (CKA).
  • Confluent Certified Administrator for Apache Kafka (CCAAK).
  • Elastic Certified Engineer / Elasticsearch Administrator.
  • Redis Certified Administrator or Redis Certified Developer.
What’s in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable.
  • Leaders who support your development through coaching and managing opportunities.
  • Ability to make a difference and lasting impact.
  • Work in a dynamic, collaborative, progressive, and high‑performing team.
  • A world‑class training program in financial services.
  • Opportunities to do challenging work.

#TECHPJ

Job Skills

Agile Methodology, Cloud Computing, Cloud Computing Architecture, Cloud Platform, Infrastructure As Code (IaC), IT Systems Integration, Performance Measurement, Requirements Analysis, Software Development, Systems Software

Additional Job Details

Address: RBC CENTRE, 155 WELLINGTON ST W:TORONTO

City: Toronto

Country: Canada

Work hours/week: 37.5

Employment Type: Full time

Platform: TECHNOLOGY AND OPERATIONS

Job Type: Regular

Pay Type: Salaried

Posted Date: 2026-09-17

Application Deadline: 2026-09-24

Note

Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community

Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.

Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well‑being of our clients and communities at jobs.rbc.com

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Technical Lead – AI & Platform
Staff Technical Lead – AI & Platform

RBC • Toronto

On-site
CAD 180,000 - 240,000
Total rewards program
Flexible benefits
Stock options where applicable
+2
Senior Software Developer
Senior Software Developer

RBC • Toronto

On-site
CAD 115,000 - 160,000
Coaching and learning opportunities
Collaborative team environment
Senior Software Engineer
Senior Software Engineer

RBC • Toronto

On-site
CAD 120,000 - 190,000
AI Engineer (Global Security)
AI Engineer (Global Security)

RBC • Toronto

On-site
CAD 120,000 - 160,000
Total rewards program
Flexible work-life balance
Career development and coaching
+1
Staff Engineer – Software Development, Architecture, Security & Governance
Staff Engineer – Software Development, Architecture, Security & Governance

RBC • Toronto

On-site
CAD 140,000 - 190,000
Total rewards program
Coaching & development
World-class training
Senior Java Developer - GFT TORONTO
Senior Java Developer - GFT TORONTO

RBC • Toronto

On-site
CAD 110,000 - 140,000
Total rewards program
Coaching & development
Work-life balance
+2
Principal Engineer, Cyber Technology Operations SRE (Global Security)
Principal Engineer, Cyber Technology Operations SRE (Global Security)

RBC • Vancouver

On-site
CAD 150,000 - 210,000
Total Rewards Program
Bonuses
Stock options
+1
Lead Full Stack Developer
Lead Full Stack Developer

RBC • Toronto

On-site
CAD 120,000 - 190,000
Total rewards program
Career development
Impact and contribution
+2
Senior Software Developer, Security Automation(Global Security)
Senior Software Developer, Security Automation(Global Security)

RBC • Toronto

On-site
CAD 120,000 - 170,000
Total rewards program
Coaching and development
World-class tools and training
+1
Senior Software Engineer
Senior Software Engineer

RBC • Calgary

On-site
CAD 120,000 - 180,000