Senior Site Reliability Engineer - Resilience & Automation

Scotiabank

Toronto

On-site

CAD 100,000 - 130,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Upskilling opportunities
Competitive rewards including bonuses
Dynamic collaborative environment
Community engagement programs

Job summary

A leading financial institution in Toronto is seeking a Site Reliability Engineer (SRE) to enhance operational performance and reliability. The successful candidate will be responsible for implementing service improvements using metrics from various tools and automating processes to minimize disruptions. Ideal applicants will possess 7+ years of technical experience in production support, cloud environments, and real-time data projects. The role offers a diverse and inclusion-focused workplace with competitive rewards including flexible vacation options.

Qualifications

  • 3+ years working with Real Time streaming data projects.
  • 7+ years in production support and incident management.
  • 2+ years using Apache Kafka for event management.
  • 3+ years using Splunk/Dynatrace for monitoring.

Responsibilities

  • Implement, measure, and improve service availability and performance.
  • Lead and perform Disaster Recovery exercises.
  • Conduct vulnerability assessments and recommend enhancements.

Skills

Real Time streaming data projects
Production support and troubleshooting
Apache Kafka
Splunk and/or Dynatrace
CI/CD deployment pipelines
Java troubleshooting
RESTful Services
SQL Queries
Microservices in Cloud (GCP, Azure)
UNIX shell scripting and Python
Organizational skills

Education

Post-secondary education in Computer Science, Engineering, or Mathematics

Tools

Apache Kafka
Splunk
Dynatrace
Jenkins
Gradle/Maven
Bitbucket

Job description

A leading financial institution in Toronto is seeking a Site Reliability Engineer (SRE) to enhance operational performance and reliability. The successful candidate will be responsible for implementing service improvements using metrics from various tools and automating processes to minimize disruptions. Ideal applicants will possess 7+ years of technical experience in production support, cloud environments, and real-time data projects. The role offers a diverse and inclusion-focused workplace with competitive rewards including flexible vacation options.
Get your free, confidential resume review.
or drag and drop your file here.