Senior Site Reliability Engineer

Morningstar

Toronto

On-site

CAD 90,489 - 132,711

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model

Job summary

Morningstar is seeking an experienced Site Reliability Engineer to design and operate scalable cloud infrastructure in Toronto. You will lead CI/CD pipelines, manage AWS resources with IaC, and drive reliability initiatives across distributed teams.

The role requires 5+ years in SRE/DevOps, strong Linux skills, and proficiency with monitoring, logging, and automation. A hybrid work setup is supported with onsite collaboration four days a week.

Qualifications

  • 5+ years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles.

Responsibilities

  • Design, build, and improve CI/CD pipelines to accelerate software delivery while maintaining stability and security across our platform.
  • Provision, configure, and maintain cloud infrastructure on AWS using IaC tools such as Terraform, CDK, or CloudFormation.
  • Provide on‑call technical triage and troubleshooting, driving incidents to resolution and conducting post‑incident reviews.
  • Lead cross‑team reliability initiatives, including disaster recovery planning, security compliance, and AWS resource optimization.
  • Deploy and manage containerized applications using Docker and AWS ECS/EKS, optimizing resource utilization and deployment strategies.
  • Drive automation and innovation for proactive monitoring, alerting, and continuous operational improvement using tools such as Splunk, CloudWatch, New Relic, and Harness.
  • Collaborate with software engineers and data engineers to embed SRE best practices into the development lifecycle, including SLOs, error budgets, and capacity planning.
  • Write scripts and tooling in Python, Bash, or other scripting languages to automate routine operational tasks and streamline deployments.
  • Document infrastructure architecture, deployment processes, and operational runbooks to enable transparency, consistency, and long‑term maintainability.
  • Collaborate with globally distributed teams for projects, knowledge transfer, and on‑call rotation coverage.
  • Leverage AI‑assisted development tools (e.g., GitHub Copilot, Claude Code) to accelerate engineering workflows and improve productivity.

Skills

AWS
CI/CD
Docker
Python
Bash
SRE concepts
Linux/Unix
Monitoring/Logging
AI-assisted tools
Communication

Education

Bachelor's degree in Computer Science/Engineering

Tools

Terraform
CDK
CloudFormation
Jenkins
GitHub Actions
Harness
Datadog
CloudWatch
Splunk
New Relic
ECS/EKS
Lambda

Job description

About the Team

Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations. We enable the Managed Investment Data (MID), Reference Entity Data (RED), Fixed Income, and Third Party Data, and Manager Research teams to collect, process, and deliver high-quality investment data at scale—supporting over 770,000 investments across thousands of global processes.

We design and maintain the internal tools and systems that support data collection for operational, performance, portfolio, and document data; enable automation and AI‑assisted workflows; improve analyst productivity and experience; and ensure data quality, scalability, and system stability. Our platforms are used by hundreds of analysts across the globe to process billions of data points every month.

Location

Toronto, ON (4 days onsite)

What You’ll Do
  • Design, build, and improve CI/CD pipelines to accelerate software delivery while maintaining stability and security across our platform.
  • Provision, configure, and maintain cloud infrastructure on AWS using Infrastructure as Code tools such as Terraform, CDK, or CloudFormation.
  • Provide on‑call technical triage and troubleshooting, driving incidents to resolution and conducting thorough post‑incident reviews.
  • Lead cross‑team reliability initiatives, including disaster recovery planning, security compliance, and AWS resource optimization.
  • Deploy and manage containerized applications using Docker and AWS ECS/EKS, optimizing resource utilization and deployment strategies.
  • Drive automation and innovation for proactive monitoring, alerting, and continuous operational improvement using tools such as Splunk, CloudWatch, New Relic, and Harness.
  • Collaborate with software engineers and data engineers to embed SRE best practices into the development lifecycle, including SLOs, error budgets, and capacity planning.
  • Write scripts and tooling in Python, Bash, or other scripting languages to automate routine operational tasks and streamline deployments.
  • Document infrastructure architecture, deployment processes, and operational runbooks to enable transparency, consistency, and long‑term maintainability.
  • Collaborate with globally distributed teams for projects, knowledge transfer, and on‑call rotation coverage.
  • Leverage AI‑assisted development tools (e.g., GitHub Copilot, Claude Code) to accelerate engineering workflows and improve productivity.
What We’re Looking For
  • 5+ years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles supporting production systems.
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • Strong hands‑on experience with AWS cloud services (EC2, S3, ECS/EKS, Lambda, RDS, VPC, IAM, Route 53, CloudWatch).
  • Proficiency with Infrastructure as Code tools such as Terraform, CDK, or CloudFormation.
  • Experience building and maintaining CI/CD pipelines using tools such as Jenkins, Harness, GitHub Actions, or similar platforms.
  • Strong working knowledge of Docker containers and container orchestration platforms.
  • Proficiency in scripting languages such as Python or Bash for automation and operational tooling.
  • Solid understanding of Linux/Unix system administration and networking fundamentals.
  • Experience with monitoring, logging, and alerting tools such as Splunk, New Relic, CloudWatch, or Datadog.
  • Knowledge of SRE principles, including SLIs/SLOs, error budgets, incident management, and post‑incident review processes.
  • Experience using AI‑assisted development tools (e.g., GitHub Copilot, Claude Code).
  • Excellent communication and collaboration skills, with the ability to work effectively across distributed teams and explain infrastructure decisions clearly.
Nice to Have
  • AWS certifications (e.g., Solutions Architect, DevOps Engineer, SysOps Administrator).
  • Experience designing or supporting disaster recovery and business continuity strategies.
  • Familiarity with security compliance frameworks and implementing security best practices in cloud environments.
  • Experience with serverless architectures using AWS Lambda, SAM, or the Serverless Framework.
  • Experience supporting distributed engineering teams across multiple time zones.
  • Exposure to data pipeline infrastructure or platforms used for large‑scale data processing.
  • FinOps certification or experience with cloud financial management and cost optimization practices.
Base Salary Compensation Range

$90,489.00-132,711.00

Incentive Target Percentage

12.5% Annual

Morningstar's hybrid work environment gives you the opportunity to collaborate in‑person each week as we've found that we're at our best when we're purposely together on a regular basis. In most of our locations, our hybrid work model is four days in‑office each week. A range of other benefits are also available to enhance flexibility as needs change. No matter where you are, you'll have tools and resources to engage meaningfully with your global colleagues.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Morningstar Credit Ratings, LLC • Toronto

On-site
CAD 90,000 - 133,000
Hybrid work environment
Flexible benefits
Senior Site Reliability Developer
Senior Site Reliability Developer

United States Digital Space LLC • Toronto

On-site
CAD 107,000 - 157,000
Salary transparency
In-person onboarding
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Sage Recruiting Inc. • Canada

On-site
CAD 180,000 - 200,000
Unlimited vacation
Comprehensive health and dental benefits
Senior Software Engineer (Full-Stack JavaScript)
Senior Software Engineer (Full-Stack JavaScript)

Morningstar • Toronto

Hybrid
CAD 90,000 - 133,000
Hybrid work model
4 days in‑office per week
Software Engineer, Async Platform
Software Engineer, Async Platform

United States Digital Space LLC • Toronto

Hybrid
CAD 108,000 - 135,000
Health & dental
Mental health benefits
Family building benefits
+6
Site Reliability Engineer
Site Reliability Engineer

Momentum Financial Services Group • Toronto

On-site
CAD 110,000 - 120,000
Compensation aligned with experience
Discretionary annual bonus
Comprehensive health and dental benefits
+3
Site Reliability Engineer (SRE), Cloud Operations
Site Reliability Engineer (SRE), Cloud Operations

United States Digital Space LLC • Toronto

On-site
CAD 110,000 - 140,000
Bonuses
Flexible benefits
Stock options
+1
Lead, Site Reliability Engineering (Application Support)
Lead, Site Reliability Engineering (Application Support)

OMERS • Toronto

Hybrid
CAD 86,000 - 130,000
Site Reliability Engineer - Cloud & Platform Engineering, Manulife Bank Technology
Site Reliability Engineer - Cloud & Platform Engineering, Manulife Bank Technology

Socket.dev • Southwestern Ontario

Hybrid
CAD 86,000 - 136,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iManage • Toronto

Hybrid
CAD 90,000 - 120,000
Market-competitive salary
Annual performance-based bonus
Comprehensive Health, Vision, Dental, and Life insurance
+4