Site Reliability Engineer (SRE), ServiceNow, Application Infrastructure

Open Systems Technologies

Montreal

On-site

CAD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology firm in Montreal is seeking a Site Reliability Engineer (SRE) to enhance the reliability and operations of their ServiceNow SaaS implementation. Candidates should have software development skills, experience in collaboration, and a commitment to client service. This contract position is at a mid-senior level and requires 7+ years of relevant experience. Applicants from diverse backgrounds are encouraged to apply.

Qualifications

  • 7+ years of experience required.
  • Proficient in oral and written communication.
  • Effective team collaboration and client service skills.

Responsibilities

  • Maximize availability and performance of systems.
  • Troubleshoot ServiceNow issues in a Linux environment.
  • Deliver observability metrics and ensure product reliability.
  • Participate in on-call rotation as needed.
  • Understand and document ServiceNow instances.
  • Identify and prioritize technical debt.
  • Provide feedback on SRE policies and procedures.

Skills

Software development skills in Python
ServiceNow administration or development experience
Strong communication skills
Team collaboration
Client service commitment
Ability to respond to technical emergencies
Open to on-call rotation

Job description

Site Reliability Engineer (SRE), ServiceNow, Application Infrastructure

2 days ago Be among the first 25 applicants

The Application Infrastructure (AI) department is seeking a Site Reliability Engineer (SRE) to help drive the reliability engineering, operations and customer support services for Morgan Stanley's ServiceNow SaaS implementation. Reporting to a Site Reliability Engineering & Operations Lead.

This role requires delivering a range of SRE practices within a global community of other SREs. This means teaming up with colleagues to deliver reliable, resilient systems without wasteful operational effort.

SRE practices include task optimization and automation, prioritizing technical debt, observability and monitoring dashboards, capacity management, incident response, and problem elimination.

This position specializes in ServiceNow Software as a Service which provides a suite of IT service management capabilities and is integrated with many products such as chatbot technology, on‑call escalation incident management, and a range of other on‑premises infrastructure (including SQL databases, APIs, and web infrastructure). Despite the focus on value‑add development and process delivery, this is also a production‑side, operational role requiring participation in an on‑call rotation from time to time.

Successful candidates for SRE roles in Application Infrastructure have so far come from a variety of backgrounds; maybe a developer today looking to evolve site reliability as a practice, or an infrastructure specialist with an interest in reliability and resilience principles, or a strong system admin who enjoys troubleshooting along with some task automation experience.

Prior experience in the financial services industry is not required, and we welcome candidates from all industries and backgrounds to apply.

Responsibilities
  • Delivery of improvements that will maximize the availability and performance of supported systems through optimized and automated operational tasks, collaborating on the development of operational tools, ongoing problem management, and architecture reviews with colleagues.
  • Troubleshooting ServiceNow issues, and also some on‑premise capabilities in a Linux environment from time to time, collaborating with others to get to the bottom of issues, and agreeing on lasting improvements that can be made.
  • Exploring and delivering observability including metrics, logging, tracing and alerting that can define and measure the target reliability of a product.
  • Being dependable and responsive during agreed hours, like when part of the on‑call rotation with the rest of the global team (with a time‑off in lieu system).
  • A commitment to understanding the Firm's ServiceNow instances and related dependencies, contributing to their documentation.
  • Identification and prioritization of technical debt that can impact client satisfaction or operational efficiency.
  • Give feedback on policy and procedures related to the delivery of SRE and operational practices with a view to continually making the Firm safer and more efficient.
Skills Required
  • The ideal candidate would have at least one of: Software development skills in one or more programming language, e.g. Python, ServiceNow administration or development experience.
  • 7+ years of experience
  • Proficient oral and written communication skills
  • Establishing warm, effective relationships with colleagues to collaborate on successful delivery
  • A dependable team worker with demonstrated commitment to client service
  • Ability to respond appropriately during occasional technical emergencies, like outages.
  • Open to work in on‑call rotation
  • ServiceNow administration or development experience, although this can be acquired by the successful candidate via on‑the‑job learning and training.
Seniority level

Mid‑Senior level

Employment type

Contract

Job function

Staffing and Recruiting

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE), ServiceNow, Application Infrastructure
Site Reliability Engineer (SRE), ServiceNow, Application Infrastructure

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 90,000 - 120,000
SRE: ServiceNow Infra Reliability & Automation (Contract)
SRE: ServiceNow Infra Reliability & Automation (Contract)

Open Systems Technologies • Montreal

On-site
CAD 90,000 - 120,000
SRE for ServiceNow: Reliability, Observability & Automation
SRE for ServiceNow: Reliability, Observability & Automation

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 90,000 - 120,000
Site Reliability Engineer/ServiceNow SaaS
Site Reliability Engineer/ServiceNow SaaS

NTT DATA North America • Montreal

Hybrid
CAD 90,000 - 120,000
SRE for ServiceNow SaaS & Observability
SRE for ServiceNow SaaS & Observability

NTT DATA North America • Montreal

Hybrid
CAD 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • Montreal (administrative region)

Hybrid
CAD 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

ApTask • Montreal

On-site
CAD 125,000 - 250,000
Manager, Network Reliability and Resiliency
Manager, Network Reliability and Resiliency

Servicenow • Toronto

On-site
CAD 126,000 - 220,000
Health plans
RRSP plan with company match
ESPP
+3
Manager, Site Reliability Engineering (SRE)
Manager, Site Reliability Engineering (SRE)

Quantum Technology Recruiting Inc. (QTR) • Toronto

On-site
CAD 155,000 - 165,000
Site Reliability Engineer (SRE) – Observability
Site Reliability Engineer (SRE) – Observability

Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! • Toronto

Hybrid
CAD 75,000 - 95,000