Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)

CoSourcing Partners - Enterprise-AI and IT Services Company

Chicago (IL)

Hybrid

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology services company is seeking an Edge Site Reliability Engineer to lead the design and automation of Google Distributed Cloud Edge environments. The ideal candidate will have strong networking fundamentals and hands-on experience with Kubernetes, CI/CD pipelines, and monitoring tools. You'll define automation principles, collaborate across teams for security and performance, and drive initiatives to improve reliability. This hybrid position is located in Chicago, IL and offers a full-time employment type.

Qualifications

  • Strong expertise in networking fundamentals (TCP/IP, BGP, DNS) and carrier-grade environments.
  • Hands-on experience with Kubernetes administration and CI/CD pipelines.
  • Deep experience in monitoring/observability tools such as Prometheus and Grafana.

Responsibilities

  • Define and implement automation principles for edge compute provisioning.
  • Collaborate with security teams to implement protection and secure policies.
  • Lead incident response efforts and root cause analysis.

Skills

Networking fundamentals (TCP/IP, BGP, DNS)
Kubernetes administration
CI/CD pipelines
Infrastructure as Code (Terraform)
GCP (preferred) and/or AWS
Monitoring/observability tools (Prometheus, Grafana, ELK, New Relic)
Automation and tooling initiatives
Communication and leadership skills

Job description

Recruitment Specialist @ CoSourcing Partners | AI and IT Recruitment

Location: Hybrid – Chicago, IL (preferred)

Employment Type: W2, Contract to Hire, Direct Hire

Overview

Our client is seeking a highly skilled Edge Site Reliability Engineer (Edge SRE) to lead the design, automation, and operations of Google Distributed Cloud Edge (GDCE) environments. This role combines deep expertise in cloud-native platforms, networking, and automation with a strong focus on performance, reliability, and scalability at the edge. The ideal candidate thrives in complex, distributed systems and is experienced at bridging platform engineering with application, network, and security teams.

Key Responsibilities
  • Define and implement automation principles for edge compute provisioning and application deployments.
  • Design and optimize intelligent caching, traffic steering, and edge routing strategies to reduce latency and maximize performance.
  • Collaborate with security teams to implement DDoS protection, bot mitigation, and secure TLS termination policies.
  • Develop monitoring, alerting, and observability frameworks (real-time + historical) for latency, traffic, and system health.
  • Lead incident response efforts, including root cause analysis and blameless post-mortems.
  • Partner with application, network, security, and platform teams to ensure edge systems integrate seamlessly with core infrastructure.
  • Drive automation initiatives to reduce operational toil and improve system efficiency.
  • Build tools, dashboards, and data pipelines to monitor service performance and identify bottlenecks.
  • Establish and champion KPIs that measure reliability, scalability, and overall success of the edge practice.
  • Serve as a thought leader in advancing edge operations aligned with business goals.
Skills & Qualifications
  • Strong expertise in networking fundamentals (TCP/IP, BGP, DNS) and carrier-grade environments.
  • Hands‑on experience with Kubernetes administration, CI/CD pipelines, and Infrastructure as Code (Terraform).
  • Proven background in GCP (preferred) and/or AWS cloud infrastructure.
  • Deep experience in monitoring/observability tools (Prometheus, Grafana, ELK, New Relic, etc.).
  • Demonstrated success driving automation and tooling initiatives to improve reliability and reduce toil.
  • Prior experience guiding or leading Operations/SRE teams in large‑scale, multinational environments.
  • Exceptional communication, leadership, and cross‑functional collaboration skills.
Seniority level
  • Not Applicable
Employment type
  • Full-time
Job function
  • Information Technology
Industries
  • Information Services
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)
Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)

CoSourcing Partners Inc. • Chicago (IL)

Hybrid
USD 150,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Manager of Site Reliability Engineering (SRE)
Manager of Site Reliability Engineering (SRE)

Genuine Parts Company • Alabama

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

JPS Tech Solutions • San Jose (CA)

On-site
USD 130,000 - 160,000
Senior Site Reliability Engineer NEX
Senior Site Reliability Engineer NEX

Patterson-UTI • Houston (TX)

On-site
USD 120,000 - 180,000
Edge SRE: Distributed Cloud & Edge Platform Automation
Edge SRE: Distributed Cloud & Edge Platform Automation

CoSourcing Partners - Enterprise-AI and IT Services Company • Chicago (IL)

Hybrid
USD 120,000 - 150,000
Senior Site Reliability Engineer (SRE) – eCommerce & Google Cloud Platform (GCP)
Senior Site Reliability Engineer (SRE) – eCommerce & Google Cloud Platform (GCP)

Cognizant • Phoenix (AZ)

On-site
USD 50,000 - 70,000
Medical/Dental/Vision/Life Insurance
Paid Holidays and PTO
401(k) Plan and Company Contributions
+4
Site Rel Eng III, GCP
Site Rel Eng III, GCP

Altice USA • Bethpage (NY)

On-site
USD 134,000 - 220,000
Sr. Cloud Operations Reliability Engineer (SRE)
Sr. Cloud Operations Reliability Engineer (SRE)

NextGen Healthcare • Georgia

On-site
USD 140,000 - 210,000
Senior Software Engineer, Site Reliability Engineering
Senior Software Engineer, Site Reliability Engineering

Google • Sunnyvale (CA)

On-site
USD 174,000 - 252,000