Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)

CoSourcing Partners Inc.

Chicago (IL)

Hybrid

USD 150,000 - 190,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

CoSourcing Partners Inc. is seeking an experienced Edge Site Reliability Engineer to lead the design, automation, and operations of Google Distributed Cloud Edge environments in a hybrid Chicago setting.

The role focuses on reliability, performance, and scalability at the edge, bridging platform engineering with application, network, and security teams. You will drive automation and champion edge KPIs across a multinational landscape.

Qualifications

  • Proven expertise in cloud-native platforms and edge computing.
  • Strong networking fundamentals (TCP/IP, BGP, DNS) and carrier-grade networks.
  • Hands-on Kubernetes administration and IaC (Terraform).
  • Experience with GCP and/or AWS cloud environments.
  • Experience leading SRE teams and cross-functional collaboration.

Responsibilities

  • Define and implement automation for edge compute provisioning and deployments.
  • Design caching, traffic steering, and edge routing to reduce latency.
  • Collaborate with security to implement DDoS protection and TLS termination policies.
  • Develop monitoring and observability for latency, traffic, and health.
  • Lead incident response with blameless post-mortems.
  • Partner with cross-functional teams to integrate edge with core infra.
  • Drive automation to reduce toil and improve efficiency.
  • Build dashboards and data pipelines to monitor performance.
  • Establish KPIs for reliability and scalability.
  • Serve as thought leader for edge operations aligned with goals.

Skills

Networking basics
Kubernetes admin
CI/CD pipelines
Terraform
GCP/AWS
Monitoring tools
Automation
Leadership

Tools

Prometheus
Grafana
ELK
New Relic
CI/CD tooling

Job description

  • Chicago, IL
Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)
Location: Hybrid – Chicago, IL (preferred)
Employment Type: W2, Contract to Hire, Direct Hire
Overview

Our client is seeking a highly skilled Edge Site Reliability Engineer (Edge SRE) to lead the design, automation, and operations of Google Distributed Cloud Edge (GDCE) environments. This role combines deep expertise in cloud-native platforms, networking, and automation with a strong focus on performance, reliability, and scalability at the edge. The ideal candidate thrives in complex, distributed systems and is experienced at bridging platform engineering with application, network, and security teams.

Key Responsibilities
  • Define and implement automation principles for edge compute provisioning and application deployments.
  • Design and optimize intelligent caching, traffic steering, and edge routing strategies to reduce latency and maximize performance.
  • Collaborate with security teams to implement DDoS protection, bot mitigation, and secure TLS termination policies.
  • Develop monitoring, alerting, and observability frameworks (real-time + historical) for latency, traffic, and system health.
  • Lead incident response efforts, including root cause analysis and blameless post-mortems.
  • Partner with application, network, security, and platform teams to ensure edge systems integrate seamlessly with core infrastructure.
  • Drive automation initiatives to reduce operational toil and improve system efficiency.
  • Build tools, dashboards, and data pipelines to monitor service performance and identify bottlenecks.
  • Establish and champion KPIs that measure reliability, scalability, and overall success of the edge practice.
  • Serve as a thought leader in advancing edge operations aligned with business goals.
Skills & Qualifications
  • Strong expertise in networking fundamentals (TCP/IP, BGP, DNS) and carrier-grade environments.
  • Hands-on experience with Kubernetes administration, CI/CD pipelines, and Infrastructure as Code (Terraform).
  • Proven background in GCP (preferred) and/or AWS cloud infrastructure.
  • Deep experience in monitoring/observability tools (Prometheus, Grafana, ELK, New Relic, etc.).
  • Demonstrated success driving automation and tooling initiatives to improve reliability and reduce toil.
  • Prior experience guiding or leading Operations/SRE teams in large-scale, multinational environments.
  • Exceptional communication, leadership, and cross-functional collaboration skills.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)
Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)

CoSourcing Partners - Enterprise-AI and IT Services Company • Chicago (IL)

Hybrid
USD 120,000 - 150,000
Edge SRE: Distributed Cloud & Edge Platform Automation
Edge SRE: Distributed Cloud & Edge Platform Automation

CoSourcing Partners - Enterprise-AI and IT Services Company • Chicago (IL)

Hybrid
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

JPS Tech Solutions • San Jose (CA)

On-site
USD 130,000 - 160,000
Senior Site Reliability Engineer NEX
Senior Site Reliability Engineer NEX

Patterson-UTI • Houston (TX)

On-site
USD 120,000 - 180,000
Cloud Architect
Cloud Architect

CoSourcing Partners - Enterprise-AI and IT Services Company • Chicago (IL)

Hybrid
USD 130,000 - 170,000
Lead Site Reliability Engineer - Infrastructure & DevOps
Lead Site Reliability Engineer - Infrastructure & DevOps

SRI Tech Solutions Inc. • Orlando (FL)

On-site
USD 140,000 - 190,000
Manager of Site Reliability Engineering (SRE)
Manager of Site Reliability Engineering (SRE)

Genuine Parts Company • Alabama

On-site
USD 120,000 - 150,000
Sr. Cloud Operations Reliability Engineer (SRE)
Sr. Cloud Operations Reliability Engineer (SRE)

NextGen Healthcare • Georgia

On-site
USD 140,000 - 210,000
Senior Site Reliability Engineer (SRE) - Scale & Reliability
Senior Site Reliability Engineer (SRE) - Scale & Reliability

Socket.dev • Pittsburgh

On-site
USD 174,000 - 253,000
Bonus target