Site Reliability Engineer

Optimum

Bethpage (NY)

On-site

USD 84,000 - 137,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Optimum is seeking a Site Reliability Engineer I to support and maintain enterprise virtualization and container platforms powering critical business apps. You will work with VMware, Nutanix, and Kubernetes across multiple data centers, under guidance from senior engineers, building automation and platform engineering expertise.

The role emphasizes virtualization, Kubernetes, automation, and platform operations within a large-scale enterprise environment.

Qualifications

  • Experience with VMware vSphere, ESXi, vCenter, and clustering concepts.
  • Hands-on Kubernetes experience including deployments, services, and storage.
  • Scripting and automation using PowerShell, Python, or Bash.
  • Familiarity with CI/CD tools and Git-based workflows.
  • Experience with OpenShift, Rancher, or Tanzu is a plus.

Responsibilities

  • Support VMware, Nutanix AHV, and Kubernetes environments across data centers.
  • Perform health checks and monitor cluster, host, and VM performance.
  • Assist with upgrades, patching, lifecycle management, and maintenance.
  • Support VM provisioning, migrations, and decommissioning tasks.
  • Help operate enterprise Kubernetes workloads and containerized apps.
  • Collaborate on cluster deployment, upgrades, and operations.
  • Monitor health and availability of clusters and applications.
  • Troubleshoot infrastructure and application issues in Kubernetes.

Skills

VMware vSphere knowledge
Kubernetes platforms
Automation scripting
CI/CD familiarity
OpenShift/Rancher familiarity

Tools

PowerShell
Python
Bash
Git
Terraform
Ansible

Job description

Select how often (in days) to receive an alert:

Site Reliability Engineer

Location: Bethpage, NY, US, 11714

Brand: Optimum

Requisition #: 12431

Are you looking to Optimize your life? Start your exciting path to a rewarding career today!

We are Optimum, a leader in the fast-paced world of connectivity, and we're seeking driven and enthusiastic professionals to join our team, empower lives, fuel businesses, and drive innovation. Connectivity is no longer a luxury, but a necessity. A career at Optimum means you'll be enabling progress and enhancing lives by providing reliable, high-speed connectivity solutions that keep the world connected. Our successes, now and in the future, are powered by our amazing product, a commitment to our people and culture, and the connections we make in our communities.

If you are resourceful, collaborative, and passionate about delivering consistent excellence, Optimum is for you!

Job Summary

As a Site Reliability Engineer I (Hypervisor & Kubernetes Platform), you will support and maintain the enterprise virtualization and container platforms that power critical business applications. Working alongside senior engineers, you will help ensure the reliability, availability, and performance of VMware, Nutanix, and Kubernetes environments across multiple data centers.

This role is ideal for an engineer with foundational infrastructure experience who is looking to build expertise in virtualization, Kubernetes, automation, and platform engineering within a large-scale enterprise environment.

Responsibilities
  • Support VMware vSphere, VMware Cloud Foundation (VCF), and Nutanix AHV environments.
  • Perform health checks and monitor cluster, host, and VM performance.
  • Assist with upgrades, patching, lifecycle management, and infrastructure maintenance.
  • Support VM provisioning, migrations, and decommissioning activities.
  • Support enterprise Kubernetes environments and containerized workloads.
  • Assist with cluster deployment, upgrades, and ongoing operations.
  • Monitor cluster health, node performance, and application availability.
  • Troubleshoot Kubernetes infrastructure, networking, storage, and application issues.
  • Support development teams with onboarding and operating workloads on Kubernetes platforms.
Site Reliability Engineering
  • Participate in incident response and production support activities.
  • Assist in root cause analysis (RCA) and service restoration efforts.
  • Develop and maintain operational runbooks and support documentation.
  • Contribute to monitoring improvements and service reliability initiatives.
  • Develop scripts and automation using PowerShell, Python, Bash, or similar tools.
  • Support Infrastructure-as-Code and platform automation initiatives.
  • Assist with CI/CD integrations and operational workflow automation.
  • Identify opportunities to reduce manual effort and improve platform efficiency.
Technical Requirements
  • Working knowledge of VMware vSphere, ESXi, vCenter, clusters, HA, and DRS.
  • Familiarity with Nutanix AHV and hyperconverged infrastructure concepts.
  • Understanding of virtual machine lifecycle management and infrastructure operations.
  • Working knowledge of Kubernetes and containerized applications.
  • Familiarity with core Kubernetes concepts including pods, deployments, services, ingress, namespaces, and persistent storage.
  • Experience supporting Kubernetes clusters and troubleshooting application, networking, and storage issues.
  • Exposure to Talos, OpenShift, Rancher, or similar Kubernetes platforms preferred.
Infrastructure & Networking
  • Understanding of TCP/IP, DNS, VLANs, routing, and load balancing.
  • Familiarity with SAN, NAS, HCI, and virtualization storage platforms.
  • Knowledge of infrastructure components supporting both Kubernetes and virtualized workloads.
Automation & Scripting
  • Basic proficiency with PowerShell, Python, Bash, or similar scripting languages.
  • Familiarity with Git and source control concepts.
  • Exposure to Infrastructure as Code (Terraform, Ansible, GitOps) is a plus.
Qualifications

Preferred Qualifications

  • 1-3 years of experience supporting enterprise infrastructure environments.
  • Exposure to VMware, Nutanix, or Kubernetes platforms through professional, academic, lab, or home environments.
  • Familiarity with Talos Linux, OpenShift, Rancher, Tanzu, or other Kubernetes distributions.
  • Strong analytical, troubleshooting, and problem-solving skills.

At Optimum, every action and interaction we take part in, is driven by our three Guiding Principles: Do What’s Right, Drive One Optimum, and Make It Happen. These aren’t just words, they help us build trust, create real community, and embrace new ways of thinking. Our employees are empowered to do the right thing for our customers and co-workers and to recognize and reward these behaviors when we see them. It’s all part of the bigger picture of “Be The Difference” where each employee knows they have the power to enact real change, share new ideas, and understand that learning never stop.

If you have the drive to succeed and are ready to embark on a thrilling career, seize this opportunity today, and join our winning team. Together, we'll shape the future of connectivity.

All job descriptions and required skills, qualifications and responsibilities for a particular position are subject to modification by the Company from time to time, in the Company’s discretion based on business necessity.

We are an Equal Opportunity Employer committed to recruiting, hiring and promoting qualified people of all backgrounds regardless of gender, race, color, creed, national origin, religion, age, marital status, pregnancy, physical or mental disability, sexual orientation, gender identity, military or veteran status, or any other basis protected by federal, state, or local law.

Pay is competitive and based on a number of job-related factors, including skills and experience. The starting pay rate/range at time of hire for this position in the posted location is $83,538.00 - $137,241.00 / year. The rate/range provided herein is the anticipated pay at the time of hire and does not reflect future job opportunity. We appreciate your interest in this opportunity. Applicants must be authorized to work for ANY employer in the U.S. Please note that at this time, we do not provide visa sponsorship for employment.

Nearest Major Market: Long Island Nearest Secondary Market: New York CIty

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Altice USA • Bethpage (NY)

On-site
USD 84,000 - 137,000
Site Reliability Engineer
Site Reliability Engineer

Optimum Communications Inc. • Bethpage (NY)

On-site
USD 84,000 - 137,000
Sr Engineer Site Reliability
Sr Engineer Site Reliability

Optimum • Bethpage (NY)

On-site
USD 100,000 - 143,000
Manager, Infrastructure & Platform
Manager, Infrastructure & Platform

Optimum • Bethpage (NY)

On-site
USD 134,000 - 220,000
Site Reliability Engineer III, GCP
Site Reliability Engineer III, GCP

Optimum • Bethpage (NY)

On-site
USD 134,000 - 220,000
Site Reliability Engineer - Hypervisor & Kubernetes
Site Reliability Engineer - Hypervisor & Kubernetes

Optimum Communications Inc. • Bethpage (NY)

On-site
USD 84,000 - 137,000
Manager, Infrastructure & Platform
Manager, Infrastructure & Platform

Optimum Communications Inc. • Bethpage (NY)

On-site
USD 134,000 - 220,000
Site Reliability Engineer: Kubernetes & VM Platform
Site Reliability Engineer: Kubernetes & VM Platform

Optimum • Bethpage (NY)

On-site
USD 84,000 - 137,000
Sr Engineer Site Reliability
Sr Engineer Site Reliability

Optimum Communications Inc. • Bethpage (NY)

On-site
USD 100,000 - 143,000
Site Reliability Engineer II, Load Balancing
Site Reliability Engineer II, Load Balancing

Optimum • Bethpage (NY)

On-site
USD 100,000 - 165,000