Senior Lead Engineer, Platform Operations & Observability

McKesson

Canada

On-site

CAD 131,000 - 218,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Rémunération compétitive
Travail hybride

Job summary

McKesson Canada recherche un Senior Lead Engineer, Platform Operations & Observability pour diriger la fiabilité et l’excellence opérationnelle des plateformes technologiques d’entreprise. Vous piloterez la surveillance, les incidents et les pratiques de déploiement, en apportant leadership technique et mentorat aux équipes d’ingénierie.

Vous travaillerez avec les équipes software, plateforme et sécurité pour améliorer la fiabilité, l’évolutivité et la préparation à la production, dans un cadre

Qualifications

  • Expérience pratique avec la surveillance, l'observabilité, la journalisation et les alertes dans des environnements de production.
  • Expérience en gestion d'incidents et en support de production.
  • Connaissance des CI/CD, de l'automatisation et des pipelines de livraison de logiciels.
  • Connaissance des microservices, APIs et architectures cloud.
  • Conduite des revues de production readiness et respect des exigences opérationnelles avant les releases.

Responsibilities

  • Diriger les stratégies de surveillance et d'observabilité sur les applications et plateformes d'entreprise.
  • Concevoir et optimiser les tableaux de bord, alertes et télémétrie.
  • Conduire les processus d'incidents et les post-mortems, et coordonner les résolutions.
  • Réaliser des RCA et piloter les actions préventives et correctives.
  • Collaborer avec les équipes pour améliorer la fiabilité, l'évolutivité et l'état de préparation à la production.
  • Conduire les revues de changement et les pratiques de déploiement sûr.
  • Encadrer et coacher les ingénieurs, promouvoir les meilleures pratiques d'ingénierie.
  • Influencer l'architecture, l'automatisation et les initiatives d'excellence opérationnelle.

Skills

Monitoring
Observability
Incident management
CI/CD
Cloud architectures
Kubernetes
DevOps practices
SRE
Leadership

Education

Bachelor's degree in Computer Science/Engineering or equivalent

Tools

Dynatrace
Prometheus
Dotcom Monitor
Kubernetes
Azure

Job description

McKesson, l'un des 10 premières entreprises du classement Fortune Global 500, touche à pratiquement tous les aspects des soins de santé et s'emploie à faire une réelle différence. Nous sommes reconnus pour notre capacité à offrir un savoir, des produits et des services qui rendent les soins de qualité plus accessibles et plus abordables. Chez nous, la santé, le bonheur et le bien-être de nos gens et des personnes que nous desservons sont prioritaires-et nous tiennent à cœur.

Ce que tu fais chez McKesson a de l'importance. Nous favorisons une culture où tu peux t'épanouir et avoir un impact, et où tu es encouragé à proposer de nouvelles idées. Ensemble, nous façonnons l'avenir de la santé pour nos patients, nos communautés et nos équipes. Si tu souhaites dès aujourd'hui contribuer à la santé de demain, nous aimerions avoir de tes nouvelles.

McKesson is an impact-driven, Fortune 10 company that touches virtually every aspect of healthcare. We are known for delivering insights, products, and services that make quality care more accessible and affordable. Here, we focus on the health, happiness, and well-being of you and those we serve - we care.

What you do at McKesson matters. We foster a culture where you can grow, make an impact, and are empowered to bring new ideas. Together, we thrive as we shape the future of health for patients, our communities, and our people. If you want to be part of tomorrow's health today, we want to hear from you.

About the Role

Team/Project: Canada B2C Digital Solution. Main application in the portfolio is a B2C Platform for pharmacy patients. Team of around 20 people.

McKesson is seeking a Senior Lead Engineer, Platform Operations & Observability to lead the reliability, observability, and operational excellence of enterprise healthcare technology platforms. In this role, in accordance with Application Monitoring & Observability Lead, you will drive monitoring strategies, incident management practices, change governance, and root cause analysis initiatives while helping build scalable, secure, and resilient systems.

You will collaborate with software engineering, platform, security, and operations teams to improve service reliability, automate operational processes, and establish best practices for production readiness. This position also provides technical leadership and mentorship to engineering teams while influencing reliability standards and long-term platform strategy.

What You'll Do
  • Lead monitoring and observability strategies across enterprise applications, platforms, and services.
  • Design, implement, and optimize dashboards, alerts, telemetry, logging, and performance monitoring solutions.
  • Drive incident management processes, major incident response, escalation coordination, service restoration activities and postmortem incident.
  • Conduct root cause analysis (RCA) investigations and lead corrective and preventive action planning.
  • Partner with engineering teams to improve platform reliability, resiliency, scalability, and operational readiness.
  • Lead change management reviews and promote safe deployment and release practices.
  • Ability to execute regression and validation test plans following each production deployment.
  • Provide technical leadership, coaching, and mentoring to engineers while establishing engineering best practices.
  • Influence architecture, automation, CI/CD, and operational excellence initiatives supporting enterprise platforms.
Basic Requirements
  • 7+ years of professional experience in Software Engineering, Site Reliability Engineering, Platform Engineering, DevOps, or related technical roles.
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent experience.
  • Experience supporting large-scale production environments and enterprise applications.
  • Hands-on experience with monitoring, observability, logging, alerting, and application performance monitoring tools.
  • Proven experience leading incident management and production support activities.
  • Experience performing root cause analysis and implementing preventive solutions.
  • Experience with CI/CD, automation, DevOps practices, and software delivery pipelines.
  • Experience with microservices, APIs, distributed systems, and cloud-based architectures.
  • Lead production readiness reviews and operational acceptance activities prior to major releases.
Preferred Skills / Experience
  • Experience with tools such as Dynatrace, Prometheus, Dotcom Monitor or similar observability platforms.
  • Experience with Kubernetes, containers, and cloud platforms such as Azure.
  • Knowledge of ITIL-aligned incidents, problems, and change management practices.
  • Experience defining SLAs, MTTR and service reliability metrics.
  • Demonstrated technical leadership and mentorship of engineering teams.
  • Experience operating in regulated or highly compliant environments.
  • Experience driving platform modernization and operational excellence initiatives.
  • Strong analytical, troubleshooting, continuous improvement and stakeholder communication skills.
  • Fair understanding and mastery of Ai tools (Copilot, Rovo) and Ai agents.
Travel / Work Environment / Physical Requirements
  • Hybrid, two mandatory days at the Dobrin Office (Usually on Monday and Wednesday).
  • Ability to work at a computer for extended periods and participate in virtual collaboration activities.
  • Participation in on-call support and prod deployment rotations out of business hours may be required based on organizational needs.

We are proud to offer a competitive compensation package at McKesson as part of our Total Rewards. This is determined by several factors, including performance, experience and skills, equity, regular job market evaluations, and geographical markets. The pay range shown below is aligned with McKesson's pay philosophy, and pay will always be compliant with any applicable regulations. In addition to base pay, other compensation, such as an annual bonus or long-term incentive opportunities may be offered. For more information regarding benefits at McKesson, click here.

Notre échelle salariale de base pour ce poste

Our Base Pay Range for this position

$94,400 - $157,300

McKesson is an Equal Opportunity Employer

McKesson provides equal employment opportunities to applicants and employees, without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, age, genetic information, or any other legally protected category. For additional information on McKesson's full Equal Employment Opportunity policies, visit our Equal Employment Opportunity page.

McKesson is committed to being an Equal Employment Opportunity Employer and offers opportunities to all job seekers including job seekers with disabilities. If you need a reasonable accommodation to assist with your job search or application for employment, please contact us by sending an email to (United States) Disability_Accommodation@McKesson.com or (Canada) Accessibility@mckesson.ca. Resumes or CVs submitted to this email box will not be accepted.

Join us at McKesson!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Lead Engineer, Platform Operations & Observability/Ingénieur principal, Opérations de plateforme et observabilité
Senior Lead Engineer, Platform Operations & Observability/Ingénieur principal, Opérations de plateforme et observabilité

McKesson • Montreal (administrative region)

Hybrid
CAD 131,000 - 218,000
Senior Lead Engineer, Platform Operations & Observability/Ingénieur principal, Opérations de pl[...]
Senior Lead Engineer, Platform Operations & Observability/Ingénieur principal, Opérations de pl[...]

McKesson • Montreal (administrative region)

Hybrid
CAD 131,000 - 218,000
Lead, Enterprise Operations & Integration Services
Lead, Enterprise Operations & Integration Services

McKesson • Montreal (administrative region)

On-site
CAD 163,549 - 272,535
Gestionnaire principal, Ingénierie du Data Hub et des Produits de Données /Senior Manager, Data[...]
Gestionnaire principal, Ingénierie du Data Hub et des Produits de Données /Senior Manager, Data[...]

McKesson • Montreal (administrative region)

On-site
CAD 137,000 - 228,000
Gestionnaire, opérations numériques / Manager, Digital Operations (12 month contract)
Gestionnaire, opérations numériques / Manager, Digital Operations (12 month contract)

McKesson • Montreal (administrative region)

On-site
CAD 82,000 - 138,000
Gestionnaire principal, Ingénierie du Data Hub et des Produits de Données /Senior Manager, Data Hub Engineering & Data Products
Gestionnaire principal, Ingénierie du Data Hub et des Produits de Données /Senior Manager, Data Hub Engineering & Data Products

McKesson • Montreal (administrative region)

On-site
CAD 136,000 - 229,000
Responsable des opérations, soutien et maintenance des solutions et gestion des fournisseurs/Operations Lead, Solution Support, Maintenance & Vendor Management
Responsable des opérations, soutien et maintenance des solutions et gestion des fournisseurs/Operations Lead, Solution Support, Maintenance & Vendor Management

McKesson • Montreal (administrative region)

Hybrid
CAD 116,000 - 194,000
Gestionnaire de projet / Project Manager
Gestionnaire de projet / Project Manager

McKesson • Canada

Hybrid
CAD 109,000 - 182,000
Software Quality Assurance Engineer (QA Engineer)
Software Quality Assurance Engineer (QA Engineer)

McKesson • Canada

Hybrid
CAD 108,000 - 179,000
Gestionnaire de projets, Finance / Finance Project Manager, Pricing
Gestionnaire de projets, Finance / Finance Project Manager, Pricing

McKesson • Quebec

On-site
CAD 132,000 - 220,000