Staff Engineer - Cloud Platform

Lever, Inc.

Canada

Remote

CAD 141,000 - 240,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote work within Canada
Bonus potential
Mentoring and career development
Modern tech stack

Job summary

Lever, Inc. is seeking a Staff Engineer - Cloud Platform to design, build, and operate scalable cloud infrastructure in Canada. You will drive modernization across cloud, networking, and telecom environments, and mentor engineers while delivering reliable carrier-grade services.

You will work with GKE, NATS, HAProxy, Terraform, and CI/CD tooling, influencing platform direction and engineering standards in a distributed team.

Qualifications

  • Bachelor’s degree or equivalent practical experience.
  • 8+ years designing and operating large-scale distributed systems.
  • 5+ years hands-on with public cloud platforms, preferably Google Cloud Platform.
  • Strong Kubernetes and containerized platforms expertise.
  • Deep networking knowledge: TCP/IP, routing, NAT, DNS, load balancing.
  • Experience configuring HAProxy in highly available envs.
  • Hands-on with NATS and event-driven architectures.
  • IaC with Terraform.
  • Python, Go, or Bash scripting.
  • Experience designing CI/CD pipelines and DevOps platforms.
  • Monitoring/observability and incident management knowledge.
  • Strong problem-solving and cross-team communication.

Responsibilities

  • Design, architect, and operate scalable cloud platform infrastructure with modernization.
  • Define platform reliability, observability, DR, capacity, and ops excellence.
  • Troubleshoot end-to-end issues across cloud, network, and telecom environments.
  • Design and operate NATS and HAProxy traffic management.
  • Build scalable service communication frameworks and resiliency patterns.
  • Manage Google Cloud infrastructure: GKE, Compute Engine, Cloud Storage, BigQuery, Pub/Sub, Cloud SQL, Composer/Airflow.
  • Implement IAC using Terraform and Terragrunt for provisioning and compliance.
  • Develop reusable platform services and self-service capabilities.
  • Maintain CI/CD pipelines with Jenkins, GitLab CI, GitHub Actions.
  • Establish monitoring, logging, tracing, alerting with Prometheus, Grafana, ELK/OpenSearch.
  • Lead RCAs and incident response; drive proactive prevention.
  • Collaborate with Product, Engineering, SRE, NOC, CS teams on platform initiatives.
  • Develop technical roadmaps, governance, and mentoring across cloud and telecom domains.

Skills

Cloud platform architecture
Kubernetes
CI/CD pipelines
Python/Go scripting
Networking fundamentals
Terraform/Terragrunt
Observability/monitoring
GKE/Google Cloud
HAProxy/NATS
Distributed systems

Education

Bachelor’s degree in CS/Engineering/Telecommunications

Tools

GKE
Compute Engine
Cloud Storage
BigQuery
Pub/Sub
Cloud SQL
Composer/Airflow
Terraform
Terragrunt
Jenkins/GitHub Actions
GitLab CI

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer - Cloud Platform based in Canada.

This is a senior individual contributor role focused on building and operating highly scalable, secure, and resilient cloud platforms for next-generation broadband services.
You will shape platform architecture, modernization, automation, and reliability strategies across complex cloud and telecom environments.
The role combines deep hands‑on engineering with technical leadership across cloud infrastructure, networking, messaging, and observability.
You will work extensively with Google Cloud, Kubernetes, infrastructure as code, CI/CD, and distributed systems.
Your work will help engineering teams deliver products faster while maintaining carrier‑grade reliability and operational excellence.
You will collaborate across Engineering, SRE, Network Operations, Product, and Customer Success teams in a highly distributed environment.
The position offers significant scope to influence technical direction, mentor engineers, and drive strategic platform initiatives.

Accountabilities
  • Design, architect, and operate highly available, scalable, secure, and resilient cloud platform infrastructure, while driving modernization through cloud-native technologies and automation.

  • Define and implement strategies for platform reliability, observability, disaster recovery, capacity management, and operational excellence.

  • Troubleshoot complex end-to-end service issues across cloud infrastructure, network services, broadband access environments, and customer-facing applications.

  • Design and operate NATS messaging infrastructure and HAProxy-based traffic management, including load balancing, routing, SSL termination, high availability, and performance optimization.

  • Build scalable service communication frameworks using event-driven architectures, service discovery, fault tolerance, and resiliency patterns.

  • Design and manage Google Cloud infrastructure, including GKE, Compute Engine, Cloud Storage, BigQuery, Pub/Sub, Cloud SQL, and Composer/Airflow.

  • Implement Infrastructure as Code using Terraform and Terragrunt, automating provisioning, deployment, compliance, and operational workflows.

  • Develop reusable platform services and self-service capabilities that improve engineering productivity and accelerate application delivery.

  • Design and maintain enterprise‑scale CI/CD pipelines using tools such as Jenkins, GitLab CI, and GitHub Actions.

  • Establish and enhance monitoring, logging, tracing, and alerting using technologies such as Prometheus, Grafana, ELK/OpenSearch, VictoriaMetrics/VictoriaLogs, and cloud monitoring services.

  • Lead root cause analysis and incident response for critical platform and network issues, identifying opportunities for proactive prevention.

  • Partner with Product, Engineering, SRE, Network Operations, technical support, and Customer Success teams to deliver strategic platform initiatives.

  • Develop technical roadmaps, contribute to architecture governance and design reviews, establish engineering standards, and mentor engineers across cloud, networking, reliability, and telecom domains.

Requirements
  • Bachelor’s degree in Computer Science, Engineering, Telecommunications, or a related discipline, or equivalent practical experience.

  • 8+ years of experience designing and operating large‑scale distributed systems in production environments.

  • 5+ years of hands‑on experience with public cloud platforms, preferably Google Cloud Platform.

  • Strong hands‑on expertise with Kubernetes and containerized platforms.

  • Deep understanding of networking fundamentals, including TCP/IP, routing and switching, NAT, DNS, DHCP, firewalls, VPN technologies, and load balancing.

  • Practical experience configuring, tuning, troubleshooting, and operating HAProxy in highly available environments.

  • Hands‑on experience with NATS and event‑driven architectures.

  • Experience supporting telecommunications or broadband infrastructure, including OLTs, ONTs, routers, and access‑network technologies.

  • Experience implementing Infrastructure as Code with Terraform.

  • Proficiency in Python, Go, Bash, or comparable scripting and programming languages.

  • Experience designing and operating CI/CD pipelines and DevOps platforms.

  • Strong knowledge of monitoring, telemetry, observability, incident management, and production troubleshooting.

  • Strong problem‑solving skills and the ability to investigate complex technical issues from infrastructure through application layers.

  • Ability to work effectively across engineering disciplines and communicate technical concepts clearly with both technical and non‑technical stakeholders.

  • Experience in broadband access networks, PON technologies, or service‑provider environments is an advantage.

  • Experience with Kubernetes networking, service mesh technologies, Kafka, Redis, Elasticsearch/OpenSearch, or distributed databases is beneficial.

  • Knowledge of cloud and telecom security best practices and Google Cloud professional certifications are considered advantages.

  • Previous experience in SRE, Platform Engineering, or Telecom Cloud Operations environments is a plus.

Benefits
  • Annual base salary range of CAD $141,000–$240,000, depending on geographic location and factors such as job‑related knowledge, skills, and experience.

  • Potential eligibility for a bonus as part of the overall compensation package.

  • Remote work opportunity within Canada.

  • Opportunity to work on large‑scale cloud, networking, and broadband infrastructure supporting next‑generation digital services.

  • Exposure to modern technologies including Google Cloud, Kubernetes, Terraform, NATS, HAProxy, CI/CD, and advanced observability platforms.

  • Opportunities to influence technical strategy, architecture, engineering standards, and platform modernization.

  • Career development and mentoring opportunities within a multidisciplinary engineering environment.

  • Collaboration with distributed teams across cloud engineering, SRE, networking, operations, product, and customer‑facing functions.

  • Comprehensive benefits package; specific benefits details are shared during the recruitment process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cloud Platform and Infrastructure Specialist, Google Cloud
Cloud Platform and Infrastructure Specialist, Google Cloud

Google Canada • Toronto

On-site
CAD 170,000 - 174,000
Staff Engineer - Cloud Platform
Staff Engineer - Cloud Platform

Calix • Canada

On-site
CAD 141,000 - 240,000
Bonus eligibility
Canada location support
Solution Architect III, Platform, Google Cloud
Solution Architect III, Platform, Google Cloud

Google • Toronto

On-site
CAD 194,000 - 199,000
Equity
Benefits
Top Customer Solutions Developer II, Networking, Google Cloud (Mandarin)
Top Customer Solutions Developer II, Networking, Google Cloud (Mandarin)

Google • Southwestern Ontario

On-site
CAD 170,000 - 174,000
Solution Architect III, Platform, FSI, Google Cloud
Solution Architect III, Platform, FSI, Google Cloud

Google • Toronto

On-site
CAD 194,000 - 199,000
Staff Systems Developer Manager, AlphaNet Core SRE
Staff Systems Developer Manager, AlphaNet Core SRE

Google Inc. • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Equity
Benefits package
Bonus target 20%
Top Customer Solutions Developer II, Networking, Google Cloud (Mandarin)
Top Customer Solutions Developer II, Networking, Google Cloud (Mandarin)

Google Canada • Southwestern Ontario

On-site
CAD 170,000 - 174,000
Solution Architect III, Platform, Google Cloud
Solution Architect III, Platform, Google Cloud

Google • Toronto

On-site
CAD 194,000 - 198,000
Staff Systems Developer Manager, AlphaNet Core SRE
Staff Systems Developer Manager, AlphaNet Core SRE

Google • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Platform Solution Architect III, Google Cloud
Platform Solution Architect III, Google Cloud

Google Inc. • Vancouver, Calgary

On-site
CAD 194,000 - 198,000