Senior SRE, Managed Gateways

KONG

Toronto

On-site

CAD 120,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Kong is seeking a Senior Site Reliability Engineer for Managed Gateways in a multi-cloud environment. You will own production reliability across AWS, GCP, and Azure, and be the technical lead for enterprise deployments.

You’ll guide a high-performing team, build scalable cloud-native systems with Kubernetes and Golang, and drive incident response, SLOs, and automation to reduce toil while collaborating with Product, Professional Services, and Customer Success.

Qualifications

  • Extensive SRE experience on highly available distributed systems.
  • Kubernetes and multi-cloud architectures (AWS, GCP, Azure).
  • Golang or similar language for automation and tooling.
  • CI/CD pipelines and infrastructure-as-code (Terraform, Ansible).
  • Monitoring/observability with Prometheus, Grafana or Datadog.
  • Experience with API gateways or related network infra is valued.

Responsibilities

  • Own production reliability for Managed Gateways across multi-cloud deployments.
  • Architect scalable, fault-tolerant cloud-native systems using Kubernetes and Golang.
  • Own monitoring, incident response, blameless post-mortems, and SLO/SLI reporting.
  • Lead automation and self-service tooling to streamline deployments and ops.

Skills

Site Reliability Engineering
Kubernetes
Cloud-native
Golang
CI/CD
IaC (Terraform/Ansible)
Monitoring/Logging
Multi-cloud (AWS/GCP/Azure)

Tools

Terraform
Ansible
Prometheus
Grafana
ELK
Datadog

Job description

Are you ready to unlock intelligence?

Senior SRE, Managed Gateways
About the Role:

Kong's Managed Gateways is the fastest-growing product in the Kong portfolio, a SaaS offering with ARR growing multifold. As a Senior Site Reliability Engineer focused on Managed Gateways, you'll be instrumental in architecting and maintaining the resilient, scalable infrastructure that powers Kong's mission-critical managed services — and you'll be the technical face of that product for the enterprise customers who depend on it, directly ensuring the reliability and performance that lets them build the next generation of connected applications at global scale.

What You’ll Do:

Cloud Gateways is a complex, multi-cloud problem — supported across AWS, GCP, and Azure — and our enterprise customers run large, often unique topologies. This role carries two equally deep engineering mandates: owning production reliability for the platform, and acting as the senior technical authority who takes an enterprise customer from kickoff to a fully successful, live implementation.

Platform & Reliability Engineering

  • Lead, mentor, and inspire a high-performing team of Site Reliability Engineers dedicated to Kong's Managed Gateway offerings.

  • Architect and implement robust, scalable, and fault-tolerant cloud-native systems using technologies like Kubernetes, Golang, and major cloud providers.

  • Own the end-to-end operational lifecycle, from proactive monitoring and alerting to incident response and blameless post-mortems, ensuring continuous service improvement.

  • Drive a culture of developer delight by implementing automation, self-service tooling, and streamlined workflows for deploying and managing API gateways.

  • Define, track, and report on key SLOs and SLIs to ensure optimal performance and reliability of Managed Gateways.

  • Champion technical debt prevention and advocate for architectural best practices that enhance system resilience and reduce operational toil.

  • Collaborate cross-functionally with Product, engineering, and Customer Success to influence roadmap decisions and ensure operational readiness for new features.

Enterprise Implementation Engineering

  • Partner directly with enterprise customers — working alongside Product leadership, Professional Services, and Customer Success — to drive end-to-end onboarding and implementation of Cloud Gateways, and productize recurring implementation patterns into repeatable playbooks and platform capabilities.

  • Bring deep, cross-cloud breadth (AWS, GCP, Azure) to handle unique customer topologies and turn complex setups into successful, production-ready deployments.

  • Serve as the technical owner of the customer relationship through implementation, primarily supporting our North America customer base, and be the escalation point Customer Success leans on for technically complex accounts.

  • Feed real-world implementation patterns and customer constraints back to Product to further contribute the roadmap.

What You’ll Bring
The Toolkit
  • Extensive experience as a Site Reliability Engineer, focusing on highly available and distributed systems.

  • Deep expertise with Kubernetes and cloud-native architectures, preferably across multiple public cloud providers (AWS, GCP, Azure).

  • Strong proficiency in Golang or similar modern programming languages for automation and tool development.

  • Proven track record in building and maintaining CI/CD pipelines and infrastructure as code (Terraform, Ansible).

  • In-depth knowledge of monitoring, logging, and alerting systems (e.g., Prometheus, Grafana, ELK stack, Datadog).

  • Experience with managed services, API gateways, or similar network infrastructure is highly desirable.

The Kong DNA
  • You take immense ownership of your systems, treating reliability as a first-class feature.

  • You operate with a sense of urgency, especially in critical situations, and drive quick, effective resolutions.

  • You thrive in a collaborative environment, actively sharing knowledge and elevating the entire team.

  • Kong moves fast, and our team's spread across continents and time zones — plans shift mid-flight, and things don't always line up neatly. You don't need everything settled to do good work. You bring your own calm to the noise, figure things out as you go, and help the people around you do the same.

Bonus Points
  • Experience with Service Mesh technologies (e.g., Istio, Linkerd).

  • Familiarity with database administration for high-throughput systems (PostgreSQL, Cassandra).

  • Contributions to open-source SRE tools or projects.

  • Relevant cloud certifications (e.g., AWS Certified DevOps Engineer, CKA).

#LI-KC1

About Kong:

Kong Inc., a leading developer of API and AI connectivity technologies, is building the infrastructure that powers the agentic era. Trusted by the Fortune 500 and startups alike, Kong's unified API and AI platform, Kong Konnect, enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI models. For more information, visit www.konghq.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, Kong Konnect
Senior Site Reliability Engineer, Kong Konnect

Kong Inc. • Toronto

On-site
CAD 100,000 - 130,000
Senior Software Engineer, Konnect Core Platform
Senior Software Engineer, Konnect Core Platform

Cacheflow • Toronto

On-site
CAD 140,000 - 180,000
Software Engineer 2, Konnect Service Catalog
Software Engineer 2, Konnect Service Catalog

Worky • Toronto

On-site
CAD 90,000 - 130,000
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes

Worky • Montreal (administrative region)

On-site
CAD 120,000 - 170,000
Laptop
Flexible work arrangements
Professional development and training
Senior Software Engineer, Konnect Admin/Billing
Senior Software Engineer, Konnect Admin/Billing

Kong • Toronto

On-site
CAD 145,000 - 165,000
Senior Software Engineer, Konnect Core Platform
Senior Software Engineer, Konnect Core Platform

Jobtailor • Toronto

On-site
CAD 120,000 - 180,000
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind Americas • Montreal (administrative region)

On-site
CAD 110,000 - 170,000
Competitive salary
Laptop provided
Professional development
+2
Senior Program Manager, Security Engineering
Senior Program Manager, Security Engineering

Kong • Toronto

Hybrid
CAD 100,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iManage • Toronto

Hybrid
CAD 90,000 - 120,000
Market-competitive salary
Annual performance-based bonus
Comprehensive Health, Vision, Dental, and Life insurance
+4
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Orion Innovation • British Columbia

On-site
CAD 100,000 - 130,000