Senior Manager, Site Reliability Engineering

Finastra

Mississauga

Hybrid

CAD 110,000 - 145,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid working arrangements
Employee Assistance Program
Wellbeing initiatives

Job summary

Finastra is seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of containerized applications. You will bridge development and operations, troubleshooting complex issues, automating support tasks, and managing deployments in a DevOps pipeline.

You will lead incident response, root-cause analysis, and automated health checks across staging and production environments, driving permanent improvements and scalable operations.

Qualifications

  • Experience in application support, production support, or systems engineering.
  • Hands-on experience troubleshooting containerized applications running in Docker and Kubernetes.
  • Proficiency in Bash and Python to automate troubleshooting and support workflows.

Responsibilities

  • DevOps Support: Create, manage, and support pipelines for deploying applications.
  • Application Troubleshooting: Diagnose runtime errors, connectivity issues, and performance bottlenecks.
  • Incident Management: Respond to production alerts, perform RCA, implement fixes.
  • Change Management: Create and follow change records for auditing.
  • Deployment Support: Validate deployments across staging and production via automated pipelines.
  • Automate Operations: Write scripts to automate health checks, log rotation, and recovery.
  • Monitoring & Alerting: Configure dashboards and alerts to detect anomalies.

Skills

Docker
Kubernetes
Bash
Python
Java apps
Log analysis
Linux
CI/CD
Grafana/Loki
AZ/Cloud ops
Disaster Recovery

Tools

Jenkins
GitHub CI
ArgoCD
AKS
Flux
Ansible
Azure CI/CD
Github
Kafka
ElasticSearch/OpenSearch
IBM MQ
DNS
Networking
Oracle DB
Redhat AMQ
Grafana

Job description

Who are we?At Finastra, we’re a global leader in financial services software, dedicated to expanding access to financial services and shaping what’s next for the industry. Our technology powers mission‑critical solutions across Lending, Payments and Universal Banking, supporting over 7,000 customers, including 80% of the world’s top 50 banks, in more than 110 countries.What will you contribute?We are seeking a Senior Manager, Site Reliability Engineering to ensure the availability, performance, and reliability of our containerized business applications. You will bridge the gap between development and operations by troubleshooting complex application issues, automating routine support tasks, and managing deployments. This position will require strong DevOps pipeline development to ensure seamless application deployment and management.Responsibilities & Deliverables:DevOps Support: Create, manage, and support pipelines that the application support teams will be utilizing to deploy applications.Application Troubleshooting: Diagnose complex runtime errors, connectivity issues, and performance bottlenecks in containerized and non-containerized applications.Incident Management: Respond to production alerts, lead root-cause analysis (RCA), and implement permanent fixes.Change Management: Understand and create change records. Follow the change management process for auditing purposes.Deployment Support: Manage and validate application deployments across staging and production environments using automated pipelines that you support.Automate Operations: Write scripts to automate routine health checks, log rotation, and recovery procedures. Innovate solutions to help the customer support and engineering teams.Monitoring & Alerting: Configure application dashboards and alerts to catch anomalies before they impact users.Required Skills & Experience:Experience: Proven experience in application support, production support, or systems engineering.Container Operations: Hands-on experience troubleshooting applications running in Docker and Kubernetes.Automation & Scripting: Proficiency in Bash and Python to automate troubleshooting and support workflows.Java application Experience: Ability to troubleshoot java based applications, keystore, heap settings, and tuning.Log Analysis: Strong skills using log aggregation tools like Grafana/Loki.Advanced Linux Skills: Confident navigating Linux environments, checking system logs, and analyzing network traffic. Ability to understand, diagnose, and fix system level issues.Disaster Recovery Knowledge: Ability to understand, improve, and manage DR environments and failoversPreferred Qualifications:CI/CD Familiarity: Experience using tools like Jenkins, GitHub CI, or ArgoCD to deploy applications.Database Basics: Oracle DB queries (updates, inserts, selects, deletes)Cloud Knowledge: Experience supporting applications hosted on Azure, AWS, or GCPTechnical Experience with: YAML, XML, AKS, Flux, Ansible, Azure CI/CD, Github, Kafka, ElasticSearch/OpenSearch, IBM MQ, DNS, Networking, Oracle DB, Redhat AMQ, Grafana and observabilityBonus experience: Understanding of NACHA and different payment rails including real time paymentsCompensation: 110 - 145k CADWe are proud to offer a range of incentives to our employees worldwide. These benefits are available to everyone, regardless of grade, and reflect the values we stand for:Flexibility: Enjoy unlimited vacation, subject to local regulations and business priorities. Benefit from hybrid working arrangements and inclusive policies such as paid time off for voting, bereavement, and sick leave.Well‑being: Access confidential one‑to‑one support through our Employee Assistance Program, connect with our network of Wellbeing Champions and Gather Groups, and take part in monthly events and initiatives designed to help you thrive—inside and outside of work.Health & Financial Security: Medical, life and disability insurance, retirement plans, lifestyle, and other benefits.*Sustainability: Paid time off for volunteering and donation‑matching opportunities to support causes that matter to you.Inclusion: Get involved in our inclusion communities, such as Count Me In, Culture@Finastra, Proud@Finastra, Disabilities@Finastra, and Women@Finastra—open to everyone who wants to participate and contribute.Career Development: Access online learning and accredited courses through our Skills & Career Navigator tool.Recognition: Take part in our global recognition program, Finastra Celebrates, and share your voice through regular employee surveys that help shape our culture and ways of working.Specific benefits may vary by location.At Finastra, each individual is unique—bringing their own ideas, perspectives, cultural backgrounds, and experiences. We learn from one another, value what makes us different, and create an environment where everyone feels included, supported, and able to be their authentic selves.Be unique. Be exceptional. Help us make a difference at Finastra.Finastra is committed to providing accessible employment practices that are in compliance with the Accessibility for Ontarians with Disabilities Act (AODA). We will accommodate applicants' needs upon request, throughout all stages of the recruitment process. Please inform us of the accommodation(s) that you may require. Information received related to accommodation will be addressed confidentially.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Manager, Site Reliability Engineering
Senior Manager, Site Reliability Engineering

Socket.dev • Mississauga

Hybrid
CAD 110,000 - 145,000
Hybrid work arrangement
Unlimited vacation
Medical, life & disability insurance
+3
Expert SRE
Expert SRE

Finastra • Mississauga

Hybrid
CAD 110,000 - 150,000
Unlimited vacation
Hybrid working arrangements
Paid time off for voting
+2
Site Reliability Engineer
Site Reliability Engineer

finastra • Mississauga

Hybrid
CAD 95,000 - 130,000
Unlimited vacation
Hybrid work model
Paid time off for voting
+4
Senior Manager, Site Reliability Engineering
Senior Manager, Site Reliability Engineering

CA1 Misys International Banking Systems Limited • Mississauga

Hybrid
CAD 110,000 - 145,000
Unlimited vacation
Hybrid work arrangements
Wellbeing program
+6
Site Reliability Engineer
Site Reliability Engineer

CA1 Misys International Banking Systems Limited • Mississauga

Hybrid
CAD 110,000 - 130,000
Unlimited vacation
Hybrid working
Health insurance
+4
Expert SRE
Expert SRE

CAP D+H Shared Services Corporation • Mississauga

Hybrid
CAD 110,000 - 150,000
Unlimited vacation
Hybrid work
Wellbeing program
+1
Sales Operations Analyst
Sales Operations Analyst

Finastra • Mississauga

Hybrid
CAD 63,000 - 75,000
Hybrid working arrangements
Paid time off for voting
Bereavement leave
+2
Cyber Security Operations Center Analyst
Cyber Security Operations Center Analyst

CAP D+H Shared Services Corporation • Mississauga

Hybrid
CAD 95,000 - 100,000
Unlimited vacation
Hybrid working arrangements
Employee wellness program
+1
Sales Operations Analyst
Sales Operations Analyst

Socket.dev • Mississauga

Hybrid
CAD 63,000 - 75,000
Unlimited vacation
Hybrid work
Health insurance
+3
Cyber Security Operations Center Analyst
Cyber Security Operations Center Analyst

Finastra • Mississauga

Hybrid
CAD 95,000 - 100,000
Unlimited vacation
Hybrid work
Wellbeing program
+1