Remote Site Reliability Engineer — Scale & Observability

Orkes

United States

Remote

USD 180,000 - 250,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Comprehensive health coverage
Flexible PTO
Support for personal development

Job summary

Orkes is hiring a Site Reliability Engineer to scale our cloud-native platform and keep services highly available across regions. You’ll own reliability, observe systems, and automate workflows in collaboration with product and engineering teams.

You have 5+ years in SRE/DevOps, strong Kubernetes, cloud (AWS/GCP/Azure), and hands-on tooling (Prometheus, Grafana, Datadog, OpenTelemetry, Terraform). This remote-friendly role offers high autonomy and impactful work shaping cloud orchestration and

Qualifications

  • 5+ years in SRE/DevOps/Platform Engineering or related roles.
  • Strong knowledge of modern cloud platforms (AWS, GCP, Azure).
  • Hands-on experience with Kubernetes and containerized environments.
  • Proficiency with observability and incident-management tooling.

Responsibilities

  • Own reliability, availability, and performance of production systems in cloud environments.
  • Define and monitor SLIs/SLOs and manage error budgets.
  • Lead incident response including detection, triage, and postmortems.
  • Improve observability with logging, monitoring, alerts, and dashboards.
  • Automate operational workflows to reduce manual toil.
  • Collaborate with engineering to improve resiliency and scalability.
  • Assist with capacity planning and performance tuning.
  • Develop internal tooling, runbooks, and best practices.
  • Support Kubernetes-based infrastructure and distributed systems at scale.
  • Act as escalation point for complex production issues.

Skills

Site Reliability Engineering
DevOps
Platform Engineering
Distributed systems
Incident management
Observability
Cloud platforms
Kubernetes
CI/CD
Troubleshooting

Tools

Kubernetes
Terraform
Prometheus
Grafana
Datadog
OpenTelemetry
ELK

Job description

Orkes is hiring a Site Reliability Engineer to scale our cloud-native platform and keep services highly available across regions. You’ll own reliability, observe systems, and automate workflows in collaboration with product and engineering teams.

You have 5+ years in SRE/DevOps, strong Kubernetes, cloud (AWS/GCP/Azure), and hands-on tooling (Prometheus, Grafana, Datadog, OpenTelemetry, Terraform). This remote-friendly role offers high autonomy and impactful work shaping cloud orchestration and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Site Reliability Engineer - Observability Expert
Remote Site Reliability Engineer - Observability Expert

Codevertex Innovations • Northern (KY)

Hybrid
USD 140,000 - 195,000
Senior Site Reliability Engineer — Remote, AWS & Observability
Senior Site Reliability Engineer — Remote, AWS & Observability

Prove • United States

Hybrid
USD 140,000 - 190,000
Wellbeing reimbursement
401k Match
Parental Leave Policy
+5
Remote Site Reliability Engineer - Cloud & Observability
Remote Site Reliability Engineer - Cloud & Observability

OneStream Software • Birmingham (MI)

On-site
USD 114,000 - 148,000
Vision insurance
Medical insurance
Life insurance
+2
Observability & SRE Engineer — Cloud, Kubernetes, & Automation
Observability & SRE Engineer — Cloud, Kubernetes, & Automation

Ontrac Solutions • Arizona

Hybrid
USD 140,000 - 190,000
Verification cost reimbursement
Senior Observability & SRE Engineer (GCP/Kubernetes)
Senior Observability & SRE Engineer (GCP/Kubernetes)

Ontrac Solutions • New York (NY)

On-site
USD 120,000 - 190,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Remote Site Reliability Engineer — Scale & Automate
Remote Site Reliability Engineer — Scale & Automate

Bright Vision Technologies • Sterling (VA)

On-site
USD 100,000 - 150,000
Observability Engineer / Site Reliability Engineer
Observability Engineer / Site Reliability Engineer

Ontrac Solutions • New York (NY)

On-site
USD 120,000 - 190,000
Senior Site Reliability Engineer – Scale & Observability
Senior Site Reliability Engineer – Scale & Observability

Inspire Brands, Inc. • Atlanta (GA)

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

BlueSky Resource Solutions • Duluth (GA)

On-site
USD 120,000 - 180,000