Sr Engineer Site Reliability

Optimum Communications Inc.

Plano (TX)

On-site

USD 100,000 - 143,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Optimum Communications Inc. is seeking a hands-on Site Reliability Engineer III to lead long-term management and stabilization of a Hybrid Cloud infrastructure, spanning GCP and on-prem Unix/Linux environments.

The role focuses on improving reliability through automation, observability, and platform engineering, with a strong emphasis on scaling and securing cloud-native platforms.

Qualifications

  • 8+ years of experience in SRE, Platform Engineering, Cloud Infrastructure.
  • 5+ years operating production workloads on Google Cloud Platform.
  • Strong expertise in GCP, GKE, Kubernetes, Terraform, Cloud Networking, Monitoring/Observability, and Infrastructure as Code.
  • Experience with NCC, Interconnect, Cloud VPN, Cloud Router, Shared VPC, and hybrid cloud networking.
  • Proven track record supporting highly available production environments.
  • Experience with SLOs, incident management, automation, and operational excellence.
  • Proficiency in Python, Bash, or similar scripting languages.
  • Strong communication and collaboration skills.

Responsibilities

  • Design and operate enterprise-scale GCP infrastructure and platforms.
  • Build and manage solutions using GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, and Cloud Monitoring.
  • Architect and support cloud networking, including NCC, Interconnect, Cloud VPN, Cloud Router, Shared VPC, and Private Service Connect.
  • Develop Infrastructure as Code using Terraform and automate deployment and workflows through CI/CD pipelines.
  • Define and manage SLOs, improve observability, and reduce toil through automation.
  • Implement cloud security, IAM, governance, and operational best practices.
  • Partner with application, network, and security teams to deliver reliable cloud-native platforms.
  • Mentor engineers and help establish network and SRE standards across the organization.

Job description

Job Summary

As a Site Reliability Engineer III, you will be a primary driver in the long-term management and stabilization of our Hybrid Cloud infrastructure. We maintain a permanent dual-hosting strategy, operating both Google Cloud Platform (GCP) and mission-critical On-Premises Unix/Linux footprint. You will bridge the gap between physical hardware and modern cloud-native operations, applying software engineering principles to ensure our systems are scalable, secure, and predictable across all platforms.

The Mission: Hybrid Reliability & Stabilization

The ideal candidate brings deep expertise in GCP, cloud networking, Infrastructure as Code, observability, and incident management, along with a passion for improving reliability through automation and engineering excellence.

Responsibilities
  • Design and operate enterprise-scale GCP infrastructure and platforms.
  • Build and manage solutions using GKE, Cloud Run, Compute Engine, Cloud SQL, Pub/Sub, Cloud Storage, and Cloud Monitoring.
  • Architect and support cloud networking, including Network Connectivity Center (NCC), Interconnect, Cloud VPN, Cloud Router, Shared VPC, and Private Service Connect.
  • Develop Infrastructure as Code using Terraform and automate deployment and operational workflows through CI/CD pipelines.
  • Define and manage SLOs, improve observability, and reduce operational toil through automation.
  • Implement cloud security, IAM, governance, and operational best practices.
  • Partner with application, network, and security teams to deliver reliable cloud-native platforms.
  • Mentor engineers and help establish network and SRE standards across the organization.
Qualifications
  • 8+ years of experience in SRE, Platform Engineering, Cloud Infrastructure, Network Engineering.
  • 5+ years of hands‑on experience operating production workloads on Google Cloud Platform.
  • Strong expertise in:
    • GCP and GKE
    • Kubernetes
    • Terraform
    • Cloud Networking
    • Monitoring and Observability
    • Infrastructure as Code
  • Experience with NCC, Interconnect, Cloud VPN, Cloud Router, Shared VPC, and hybrid cloud networking.
  • Proven experience supporting highly available production environments.
  • Experience with SLOs, incident management, automation, and operational excellence.
  • Proficiency in Python, Bash, or similar scripting languages.
  • Strong communication and collaboration skills.
Preferred Qualifications
  • Google Cloud Professional Cloud Network Engineer.
  • Google Cloud Professional Cloud Architect.
  • Google Cloud Professional Cloud DevOps Engineer.
  • Experience with GitOps, FinOps, AI/ML workloads, and hybrid cloud platforms.
What You'll Bring

You are a hands‑on engineer with deep GCP and cloud networking expertise who enjoys solving complex operational challenges through automation, platform engineering, and reliability‑focused design. You combine technical depth with leadership to build resilient, scalable cloud platforms that enable teams to deliver software with confidence.

At Optimum, we're fueled by our four core pillars: Taking Ownership, Upholding Transparency, Creating Community, and Demonstrating Expertise. Our commitment to empowering employees to take responsibility and embrace proactive problem‑solving underpins Taking Ownership. Upholding Transparency is at the core of our culture, with open and honest communication fostering trust among our dedicated team and loyal customers. Creating Community is more than a goal; it's our daily commitment to fostering an environment of collaboration, innovation, and positivity. Demonstrating expertise is a promise we uphold through continuous learning and engagement with our customers to consistently deliver top‑quality products and services. These pillars not only shape our culture but define Optimum as a place of excellence, trustworthiness, and thriving community, and we invite you to be a part of our journey.

All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company's discretion based on business necessity.

We are an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, creed, national origin, religion, age, disability, sex, sexual orientation, gender identity or protected veteran status, or any other basis protected by applicable federal, state, or local law. The Company provides reasonable accommodations upon request in accordance with applicable requirements.

Optimum collects personal information about its applicants for employment that may include personal identifiers, professional or employment related information, photos, education information and/or protected classifications under federal and state law. This information is collected for employment purposes, including identification, work authorization, FCRA‑compliant background screening, human resource administration and compliance with federal, state, and local law.

Applicants for employment with the Company will never be asked to provide money (even if reimbursable) as part of the job application or hiring process. Please review our Fraud FAQ for further details.

Pay is competitive and based on a number of job‑related factors, including skills and experience. The starting pay rate/range at time of hire for this position in New York is $100,246.00 - $143,208.00 / year. For other locations, please inquire with your recruiter. The rates/ranges provided herein are the anticipated pay at the time of hire, and do not reflect future job opportunity.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr Engineer Site Reliability
Sr Engineer Site Reliability

Optimum Communications Inc. • Bethpage (NY)

On-site
USD 100,000 - 143,000
Site Rel Eng III, GCP
Site Rel Eng III, GCP

Altice USA • Bethpage (NY)

On-site
USD 134,000 - 220,000
Site Rel Eng III, GCP
Site Rel Eng III, GCP

Optimum • Bethpage (NY)

On-site
USD 134,000 - 220,000
Site Rel Eng III, GCP
Site Rel Eng III, GCP

Altice USA • Plano (TX)

On-site
USD 150,000 - 230,000
Site Rel Eng III, GCP
Site Rel Eng III, GCP

Optimum • Plano (TX)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hobbsnews • Jersey City (NJ)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Access to paid time off
Annual discretionary incentives
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Charlotte (NC)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Paid time off
Discretionary incentive eligibility
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Plano (TX)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Paid time off
Access to resources and support
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Koitecc Solutions • Charlotte (NC)

On-site
USD 153,000 - 192,000
Benefits eligible
Senior Systems Engineer, Site Reliability Engineering, Google Cloud
Senior Systems Engineer, Site Reliability Engineering, Google Cloud

Google • Sunnyvale (CA)

On-site
USD 166,000 - 244,000