Remote Night-Shift SRE — Cloud Infra Reliability

Peraton

Reston (VA)

On-site

USD 104,000 - 166,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Peraton is seeking a Site Reliability Engineer (Night Shift) to join a team responsible for the reliability of production systems in AWS Commercial and AWS GovCloud environments. The role emphasizes OpenShift ROSA, observability, and automation, with a shift from 11pm to 7am EST and a remote work location.

The role partners with platform engineers, security, and developers to ensure reliable infrastructure that meets security requirements, with on-call incident management and deployment

Qualifications

  • United States citizen with ability to obtain Public Trust clearance.
  • Bachelor's Degree and 8+ years of experience (or HS diploma/equivalent with 12+ years).
  • 7+ years hands-on experience in site reliability engineering, DevOps, or production systems engineering.
  • Experience operating in AWS Commercial and AWS GovCloud, including OpenShift (ROSA) or comparable Kubernetes-based platforms.
  • Strong infrastructure-as-code experience with Terraform and Ansible/Ansible Tower.
  • Experience with CI/CD platforms GitLab and Jenkins, including reliability gating and deployment automation.
  • Proficient in Linux and Windows Server administration.
  • Experience with Dynatrace, Datadog, Splunk and OpenTelemetry.
  • Demonstrated ownership of an SLI/SLO and alerting program, including error budgets.
  • Scripting/automation in Python, Bash, PowerShell, or Go.

Responsibilities

  • Operate and maintain production infrastructure services and applications, ensuring availability and reliability, performance, security, and health.
  • Monitor services using SLIs, SLOs, dashboards, alerts, and observability tools; improve detection, diagnosis, and resolution of issues.
  • Define observability requirements with application teams and implement metrics, logs, traces, dashboards, and alerts.
  • Manage production incidents and service disruptions, including on-call response and root-cause analysis.
  • Execute releases through pipelines, including staging/production promotion, validation, rollback, and troubleshooting.
  • Manage lifecycle of deployed infrastructure: upgrades, patching, and configuration changes.
  • Assess and improve service resilience via capacity planning, testing, DR, backup, and failover.
  • Identify reliability risks and waste by analyzing incident trends and capacity data.
  • Automate operations using everything-as-code to improve consistency and efficiency.
  • Collaborate with platform teams to improve reliability and operability of the environment.

Skills

SRE experience
DevOps
Observability

Education

Bachelor's Degree

Tools

Terraform
Ansible Tower
ROSA/OpenShift
AWS GovCloud
GitLab
Jenkins
Dynatrace
Datadog
Splunk
OpenTelemetry

Job description

Peraton is seeking a Site Reliability Engineer (Night Shift) to join a team responsible for the reliability of production systems in AWS Commercial and AWS GovCloud environments. The role emphasizes OpenShift ROSA, observability, and automation, with a shift from 11pm to 7am EST and a remote work location.

The role partners with platform engineers, security, and developers to ensure reliable infrastructure that meets security requirements, with on-call incident management and deployment

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer
Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Night-Shift SRE (Remote) – AWS/ROSA & Observability
Night-Shift SRE (Remote) – AWS/ROSA & Observability

Peraton • Herndon (VA)

On-site
USD 104,000 - 166,000
Evening SRE – Remote Cloud & OpenShift Reliability Engineer
Evening SRE – Remote Cloud & OpenShift Reliability Engineer

Peraton • Reston (VA)

On-site
USD 104,000 - 166,000
Evening SRE - Remote Cloud Reliability Engineer (AWS/ROSA)
Evening SRE - Remote Cloud Reliability Engineer (AWS/ROSA)

Peraton • Herndon (VA)

On-site
USD 104,000 - 166,000
Remote Evening SRE: ROSA & AWS Observability Expert
Remote Evening SRE: ROSA & AWS Observability Expert

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
External Job Posting Title Site Reliability Engineer (SRE) – Night Shift
External Job Posting Title Site Reliability Engineer (SRE) – Night Shift

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Remote SRE II: Cloud, Data Ops & Incident Response
Remote SRE II: Cloud, Data Ops & Incident Response

Cohere Health, Inc. • Boston (MA)

Hybrid
USD 100,000 - 110,000
Fully remote
5% travel
Medical insurance
+7
Senior SRE - Cloud Infra, Automation & 24/7 Ops
Senior SRE - Cloud Infra, Automation & 24/7 Ops

Oracle • Reston (VA)

On-site
USD 85,000 - 210,000
Medical, dental, and vision insurance
Paid time off
401(k) Savings Plan
Remote SRE Lead - Incident, Reliability & Observability
Remote SRE Lead - Incident, Reliability & Observability

NightDragon Acquisition Corp. • United States

On-site
USD 260,000 - 280,000
Hybrid & Remote Work
Competitive Compensation
Equity package
Remote SRE Engineer — Cloud Reliability & Automation
Remote SRE Engineer — Cloud Reliability & Automation

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 106,500 - 177,500