Evening SRE - Remote Cloud Reliability Engineer (AWS/ROSA)

Peraton

Herndon (VA)

On-site

USD 104,000 - 166,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Peraton is seeking a Site Reliability Engineer for an evening shift to support production systems in AWS Commercial and GovCloud, with ROSA/OpenShift and Azure/GCP knowledge. Remote work is available, with EST hours from 3pm to 11pm.

You will ensure reliability, observability, and secure operations across multi-cloud environments in a fast-paced federal/commercial setting. The role emphasizes incident management, IaC, and automation, partnering with platform and application teams to improve

Qualifications

  • 7+ years hands-on experience in site reliability engineering, DevOps, or production systems engineering.
  • Hands-on experience operating in AWS Commercial and AWS GovCloud, including OpenShift (ROSA) or comparable Kubernetes-based platforms.
  • Strong infrastructure-as-code experience with Terraform and Ansible/Ansible Tower.
  • Experience with CI/CD platforms GitLab and Jenkins, including reliability gating and deployment automation.
  • Proficient in Linux and Windows Server administration.
  • Experience with enterprise observability tools such as Dynatrace, Datadog, Splunk and Open Telemetry.
  • Demonstrated ownership of an SLI/SLO and alerting program, including error budgets, alert rationalization, and noise reduction.
  • Scripting/automation proficiency in Python, Bash, PowerShell, or Go.
  • Experience operating in federal or regulated environments (FISMA, FedRAMP, NIST 800-53).

Responsibilities

  • Operate and maintain production infrastructure services and applications, ensuring availability, reliability, performance, security, and health.
  • Monitor services using SLIs, SLOs, dashboards, alerts, and observability tools; improve detection, diagnosis, and resolution of issues.
  • Define observability requirements with application teams and implement metrics, logs, traces, dashboards, and alerts.
  • Manage production incidents and on-call response, troubleshooting, service restoration, root-cause analysis, and post-incident actions.
  • Release applications and infrastructure through pipelines with staging and production promotions, validation, rollback, and troubleshooting.
  • Manage lifecycle of deployed infrastructure including upgrades, patches, configuration changes, maintenance, and refreshes.
  • Improve service resilience via capacity planning, testing, disaster recovery, backup, failover, and recovery testing.
  • Identify reliability risks using metrics, incidents, and capacity data to prioritize improvements.
  • Automate operational activities using a code-first approach to improve consistency and efficiency.
  • Collaborate with platform engineers and developers to identify requirements and improve environment reliability.

Skills

Terraform
Ansible
GitLab
Jenkins
Python
Bash
PowerShell
Go
Linux administration
Windows Server administration

Education

Bachelor's Degree or higher

Tools

OpenShift ROSA
Kubernetes
Dynatrace
Datadog
Splunk
OpenTelemetry

Job description

Peraton is seeking a Site Reliability Engineer for an evening shift to support production systems in AWS Commercial and GovCloud, with ROSA/OpenShift and Azure/GCP knowledge. Remote work is available, with EST hours from 3pm to 11pm.

You will ensure reliability, observability, and secure operations across multi-cloud environments in a fast-paced federal/commercial setting. The role emphasizes incident management, IaC, and automation, partnering with platform and application teams to improve

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Evening SRE – Remote Cloud & OpenShift Reliability Engineer
Evening SRE – Remote Cloud & OpenShift Reliability Engineer

Peraton • Reston (VA)

On-site
USD 104,000 - 166,000
Remote Evening SRE: ROSA & AWS Observability Expert
Remote Evening SRE: ROSA & AWS Observability Expert

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer
Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Remote Night-Shift SRE — Cloud Infra Reliability
Remote Night-Shift SRE — Cloud Infra Reliability

Peraton • Reston (VA)

On-site
USD 104,000 - 166,000
Night-Shift SRE (Remote) – AWS/ROSA & Observability
Night-Shift SRE (Remote) – AWS/ROSA & Observability

Peraton • Herndon (VA)

On-site
USD 104,000 - 166,000
Senior Cloud Reliability Architect - Remote
Senior Cloud Reliability Architect - Remote

Peraton • Northern (KY)

Hybrid
USD 112,000 - 179,000
External Job Posting Title Site Reliability Engineer (SRE) – Night Shift
External Job Posting Title Site Reliability Engineer (SRE) – Night Shift

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
External Job Posting Title Site Reliability Engineer (SRE) – Evening Shift
External Job Posting Title Site Reliability Engineer (SRE) – Evening Shift

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Remote Lead Cloud Architect – Reliability & Automation
Remote Lead Cloud Architect – Reliability & Automation

Peraton • Herndon (VA)

On-site
USD 112,000 - 179,000
Remote Lead Cloud Architect - Reliability & Automation
Remote Lead Cloud Architect - Reliability & Automation

Peraton • Reston (VA)

On-site
USD 112,000 - 179,000