Principal Observability Engineer

ISG Search Inc

Toronto

On-site

CAD 150,000 - 190,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Azure or GCP experience
Red Hat technologies
Virtualization
Backup capabilities
Identity management experience
Certifications (Splunk/AWS/Kubernetes)

Job summary

ISG Search Inc. seeks a Principal Observability Engineer to lead the architecture, design, and implementation of enterprise observability across cloud and Kubernetes environments.

You will build and improve dashboards, alerts, and monitoring using Splunk, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch, while guiding teams and driving reliability. The candidate brings 10+ years in platform/SRE roles, deep AWS/Kubernetes expertise, and IaC tooling such as Terraform and Ansible, plus strong

Qualifications

  • 10+ years of experience in Platform Engineering, SRE, Systems Engineering, or related infrastructure discipline.
  • Hands-on with Splunk and modern observability platforms including Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
  • Deep expertise with AWS, Kubernetes/OpenShift, and infrastructure automation tools such as Terraform, Ansible, Python, or Bash.
  • Proven experience designing scalable monitoring and observability solutions within large enterprise environments.
  • Strong communication and leadership skills with the ability to collaborate across technical and business teams.

Responsibilities

  • Lead the architecture, design, and implementation of enterprise observability solutions and monitoring frameworks.
  • Develop and enhance dashboards, alerting, and monitoring capabilities using Splunk, Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
  • Design and support observability across cloud and Kubernetes-based environments, driving platform performance and reliability.
  • Develop infrastructure automation using Infrastructure as Code (IaC), scripting, and configuration management tools.
  • Provide technical leadership, establish observability best practices, and mentor engineering teams.

Skills

Observability
Leadership
Architecture
Cloud
SRE
Monitoring

Tools

Splunk
Splunk Observability
OpenTelemetry
Grafana
Prometheus
AWS CloudWatch
Terraform
Ansible
Python
Bash
Kubernetes/OpenShift

Job description

Principal Observability Engineer (BBBH11976) Toronto, Canada
Job title:

Principal Observability Engineer

Duration:

12-month contract (extension possible)

Role status:
Principal tasks and responsibilities include:
  • Lead the architecture, design, and implementation of enterprise observability solutions and monitoring frameworks.
  • Develop and enhance dashboards, alerting, and monitoring capabilities using Splunk, Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
  • Design and support observability across cloud and Kubernetes-based environments, driving platform performance and reliability.
  • Develop infrastructure automation using Infrastructure as Code (IaC), scripting, and configuration management tools.
  • Provide technical leadership, establish observability best practices, and mentor engineering teams.
Our client:

A large global enterprise organization

Qualifications and pre-requisites:
  • 10+ years of experience in Platform Engineering, Site Reliability Engineering (SRE), Systems Engineering, or a related infrastructure discipline.
  • Strong hands-on experience with Splunk and modern observability platforms, including Splunk Observability, OpenTelemetry, Grafana, Prometheus, and AWS CloudWatch.
  • Deep expertise with AWS, Kubernetes/OpenShift, and infrastructure automation tools such as Terraform, Ansible, Python, or Bash.
  • Proven experience designing scalable monitoring and observability solutions within large enterprise environments.
  • Strong communication and leadership skills with the ability to collaborate across technical and business teams.
Additional information or perks:
  • Experience with additional cloud platforms (Azure or GCP), Red Hat technologies, virtualization, storage, backup, or identity management is considered an asset.
  • Relevant certifications such as Splunk, AWS, or Kubernetes are highly desirable.

isgSearch does not use artificial intelligence throughout the hiring process.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Engineer
Platform Engineer

LanceSoft, Inc. • Montreal (administrative region)

On-site
CAD 80,000 - 120,000
Site Reliability Engineer (SRE) – Observability
Site Reliability Engineer (SRE) – Observability

Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! • Toronto

Hybrid
CAD 75,000 - 95,000
Observability Engineer – Trading
Observability Engineer – Trading

Hunter Bond • Toronto

On-site
CAD 120,000 - 250,000
SRE Observability Engineer
SRE Observability Engineer

Tata Consultancy Services • Toronto

On-site
CAD 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Vertex Elite LLC • Ottawa

On-site
CAD 83,000 - 124,000
Senior Support Engineer
Senior Support Engineer

Tata Consultancy Services • Toronto

On-site
CAD 100,000 - 120,000
GCP Observability Engineer
GCP Observability Engineer

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 85,000 - 115,000
DevOps Engineer/ Dynatrace
DevOps Engineer/ Dynatrace

Motion Recruitment • Toronto

On-site
CAD 110,000 - 140,000
Medical, Dental, and Vision Insurance
Vacation Time
Dynatrace Observability Platform Engineer
Dynatrace Observability Platform Engineer

Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! • Mississauga

Hybrid
CAD 90,000 - 140,000
Senior Observability Architect
Senior Observability Architect

Suncor • Calgary

On-site
CAD 120,000 - 180,000
Competitive compensation with regional
Bonuses
Pension + savings with company match
+3