Senior Night-Shift SRE: Global Cloud Reliability (Remote)

Resilinc

United States

Remote

USD 140,000 - 210,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Fully remote
In-person meetups
Comprehensive benefits

Job summary

Resilinc is seeking an experienced Site Reliability Engineer (SRE) to join our remote-first team. You will own production incidents, improve system reliability, and drive automation across a modern cloud-native stack (Azure, Kubernetes, Kafka, Redis, PostgreSQL).

The role establishes dedicated India-based night-time coverage aligned with US business hours, addressing a critical support gap and offering high-impact work on globally used platforms.

Qualifications

  • 6-12 years of experience in SRE/DevOps or related roles.
  • Strong hands-on experience with Azure Cloud services.
  • Solid experience in Linux system administration.
  • Expertise in Docker and Kubernetes (deployment, scaling, troubleshooting).
  • Experience with Kafka, Redis, and PostgreSQL.
  • Working knowledge of Hadoop ecosystem (HDFS, Hadoop).
  • Experience with Cloudflare (CDN, security, DNS management).
  • Proficiency in CI/CD tools (GitHub, GitHub Actions).
  • Experience in Helm Charts and Kubernetes deployments.
  • Strong understanding of monitoring and logging tools (Grafana, etc.)

Responsibilities

  • Design, implement, and manage scalable and highly available systems on Azure Cloud.
  • Monitor system performance, troubleshoot issues, and ensure uptime and reliability.
  • Manage and optimize Kubernetes clusters and containerized workloads (Docker).
  • Build and maintain robust CI/CD pipelines using GitHub Actions and related tools.
  • Implement infrastructure as code and deployment automation using Helm Charts.
  • Work with distributed systems such as Kafka, Redis, PostgreSQL, Hadoop/HDFS.
  • Configure and manage Cloudflare for performance, security, and traffic routing.
  • Set up monitoring, alerting, and observability using tools like Grafana.
  • Collaborate with development teams to improve system reliability and deployment practices.
  • Perform root cause analysis (RCA) and implement preventive measures.
  • Ensure security best practices and compliance across infrastructure.

Skills

SRE/DevOps experience
Cloud-native systems
Linux administration
Docker & Kubernetes proficiency
CI/CD automation
Monitoring & logging

Tools

Azure Cloud
Kubernetes
Docker
Kafka
Redis
PostgreSQL
Hadoop/HDFS
Cloudflare
GitHub Actions
Helm Charts
Grafana

Job description

Resilinc is seeking an experienced Site Reliability Engineer (SRE) to join our remote-first team. You will own production incidents, improve system reliability, and drive automation across a modern cloud-native stack (Azure, Kubernetes, Kafka, Redis, PostgreSQL).

The role establishes dedicated India-based night-time coverage aligned with US business hours, addressing a critical support gap and offering high-impact work on globally used platforms.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (Night Shift)
Senior Site Reliability Engineer (Night Shift)

Resilinc • United States

Remote
USD 140,000 - 210,000
Fully remote
In-person meetups
Comprehensive benefits
Remote Senior SRE: Cloud Reliability & Automation Leader
Remote Senior SRE: Cloud Reliability & Automation Leader

Noctua Technology • United States

Remote
USD 149,000 - 202,000
Senior SRE: Observability & Cloud Reliability (Remote)
Senior SRE: Observability & Cloud Reliability (Remote)

Cribl • Annapolis (MD)

Remote
USD 142,000 - 195,000
Health insurance
Dental
Vision
+6
SRE Engineering Manager — Lead Reliability, Remote Flexible
SRE Engineering Manager — Lead Reliability, Remote Flexible

DevOpsChat • California (MO)

Hybrid
USD 140,000 - 220,000
Health insurance
Professional development opportunities
Remote Senior Network Reliability Engineer (SRE)
Remote Senior Network Reliability Engineer (SRE)

Gainbridge • Zionsville (IN), Northern (KY)

On-site
USD 135,000 - 190,000
Health Insurance
Dental Insurance
Vision Insurance
+4
Remote SRE II: Cloud-Native Reliability & Automation
Remote SRE II: Cloud-Native Reliability & Automation

NationsBenefits, LLC • Plantation (FL)

On-site
USD 110,000 - 160,000
Unlimited PTO
Competitive compensation & benefits
Career growth opportunities
+1
Senior Production SRE: Cloud & On-Prem Reliability
Senior Production SRE: Cloud & On-Prem Reliability

Weights & Biases • New York (NY)

On-site
USD 140,000 - 180,000
Medical Insurance
Dental Insurance
Vision Insurance
+15
Senior SRE, Remote Cloud Reliability & Observability
Senior SRE, Remote Cloud Reliability & Observability

US Diversity Job Search • Austin (TX)

On-site
USD 142,000 - 195,000
Health insurance
Dental insurance
Vision insurance
+7
Remote Senior SRE: Platform Reliability & Incidents
Remote Senior SRE: Platform Reliability & Incidents

Tamarind Intelligence • United States

Remote
USD 99,000 - 140,000
Health coverage for you and dependents
Tech spending stipend
Employee stock purchase plan
Remote SRE Manager: Lead Incidents & Reliability
Remote SRE Manager: Lead Incidents & Reliability

NationsBenefits India • United States

On-site
USD 150,000 - 210,000
Unlimited PTO
Fully remote for US-based employees
Competitive compensation
+1