SRE Architect: Cloud, Kubernetes & Observability

Symphony Communication Services

New York (NY)

On-site

USD 140,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Regional benefits
BYOB perks
Local events and team building

Job summary

Symphony Communication Services is seeking a Site Reliability Engineer to join our globally distributed, cloud-based platform. You will design, build, and operate scalable infrastructure with IaC, manage Kubernetes clusters, and champion GitOps practices.

You will automate with Python/Go and participate in 24/7 on-call duties to keep services resilient and available. The role requires hands-on experience with GCP/AWS, Terraform, Helm, ArgoCD, and strong networking, Linux, and observability

Qualifications

  • Strong IaC experience with Terraform; Terragrunt or Ansible a plus.
  • Kubernetes and Helm expertise with hands-on ops.
  • Linux administration and troubleshooting proficiency.
  • Experience with GCP and AWS cloud services.
  • Knowledge of observability tools and practices.
  • Networking fundamentals and cloud-native architectures understood.
  • Able to work independently and in teams; good communicator.
  • Capable of leading projects and multitasking in a fast-paced environment.
  • Willing to participate in a 24/7 on-call rotation.
  • Excellent written and spoken communication for an international team.

Responsibilities

  • Design, build, and maintain scalable infrastructure using IaC.
  • Manage and optimize Kubernetes clusters; use Helm and ArgoCD.
  • Administer Linux systems; ensure performance and security.
  • Work with GCP and AWS services for cloud-native solutions.
  • Implement observability with monitoring, logging, and alerts.
  • Develop automation tools using Python and Go.
  • Provide production support and participate in on-call rotation.
  • Troubleshoot complex networking issues and optimize performance.
  • Communicate across teams to align on goals and incidents.

Skills

IaC
Kubernetes
Linux
Cloud (GCP/AWS)
Observability
Python
Go
On-call
Communication
ArgoCD

Tools

Terraform
Terragrunt
Ansible
Helm
ArgoCD

Job description

Symphony Communication Services is seeking a Site Reliability Engineer to join our globally distributed, cloud-based platform. You will design, build, and operate scalable infrastructure with IaC, manage Kubernetes clusters, and champion GitOps practices.

You will automate with Python/Go and participate in 24/7 on-call duties to keep services resilient and available. The role requires hands-on experience with GCP/AWS, Terraform, Helm, ArgoCD, and strong networking, Linux, and observability

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE Engineer
SRE Engineer

ALLTECH CONSULTING SVC INC • Oregon (WI)

On-site
USD 90,000 - 120,000
Hybrid SRE — Observability & Cloud Reliability
Hybrid SRE — Observability & Cloud Reliability

JIT Global Freight Solution P Ltd • Fort Worth (TX)

On-site
USD 115,000 - 135,000
401(k) with 150% match up to 6%
Employee Share Ownership Plan
Medical, Prescription, Dental & Vision
+3
Observability Engineer / Site Reliability Engineer
Observability Engineer / Site Reliability Engineer

Ontrac Solutions • Chicago (IL)

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineer: Observability & Cloud
Senior Site Reliability Engineer: Observability & Cloud

VBeyond Corporation • Jersey City (NJ)

On-site
USD 100,000 - 260,000
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

myBridge Corporation • Austin (TX)

On-site
USD 120,000 - 160,000
SRE (Site Realiability Engineer)
SRE (Site Realiability Engineer)

STRATIS Cloud Tech Solutions INC • Arkansas

On-site
USD 100,000 - 130,000
Competitive salary and benefits
Growth and learning opportunities
Friendly and collaborative team environment
SRE Lead engineer
SRE Lead engineer

TechDigital Group • Bellevue (WA)

On-site
USD 100,000 - 130,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Site Reliability Engineer: Build Resilient, Automated Cloud
Site Reliability Engineer: Build Resilient, Automated Cloud

SRE • Puerto Rico

Hybrid
USD 120,000 - 180,000