Remote SRE Engineer - AWS ROSA & Observability

Peraton

United States

On-site

USD 104,000 - 166,000

Full time

16 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Peraton is seeking a Site Reliability Engineer to join a remote evening-shift team responsible for AWS Commercial and GovCloud production systems, ROSA/OpenShift, and Azure/GCP knowledge. The role emphasizes operating reliability, capacity planning, and everything-as-code automation to improve deployment and recovery efficiency.

The ideal candidate has 7+ years in SRE/DevOps, must be a US citizen with Public Trust clearance, and will work with a broad observability stack to ensure top-tier

Qualifications

  • Must be a U.S. Citizen with ability to obtain Public Trust clearance.
  • Bachelor's Degree or equivalent experience.
  • 7+ years hands-on experience in site reliability engineering, DevOps, or production systems engineering.
  • Experience operating in AWS Commercial and AWS GovCloud, including OpenShift ROSA or comparable Kubernetes-based platforms.
  • Strong infrastructure-as-code experience with Terraform and Ansible/Ansible Tower.
  • Experience with CI/CD platforms GitLab and Jenkins, including reliability gating and deployment automation.
  • Proficient in Linux and Windows Server administration
  • Experience with enterprise observability tools such as Dynatrace, Datadog, Splunk and Open Telemetry.
  • Demonstrated ownership of an SLI/SLO and alerting program, including error budgets, alert rationalization, and noise reduction.
  • Scripting/automation proficiency in Python, Bash, PowerShell, or Go.
  • Experience operating in federal or regulated environments (FISMA, FedRAMP, NIST 800-53).

Responsibilities

  • Operate and maintain production infrastructure services and applications, ensuring availability and reliability, performance, security, and operational health.
  • Monitor services and applications using defined SLIs, SLOs, dashboards, alerts, and other observability tools; continuously improve the detection, diagnosis, and resolution of operational issues.
  • Partner with application teams to define application observability requirements and implement appropriate metrics, logs, traces, dashboards, and alerts into the organization's observability tooling.
  • Manage production incidents and service disruptions, including on-call response, troubleshooting, service restoration, root-cause analysis, and post-incident corrective actions.
  • Execute application and infrastructure releases through established deployment pipelines, including promotion through staging and production, validation, rollback, and release-related troubleshooting.
  • Manage the operational lifecycle of deployed infrastructure, including upgrades, patching, configuration changes, maintenance, and technology refreshes.
  • Assess and improve service resilience through capacity planning, performance testing, failure-mode analysis, disaster recovery, backup, failover, and recovery testing.
  • Identify and address reliability risks and operational technical debt by using reliability metrics, incident trends, capacity data, and service health indicators to prioritize improvements.
  • Automate operational activities using an everything-as-code approach to improve consistency, repeatability, testing, deployment, recovery, and operational efficiency.
  • Collaborate with platform engineering and application teams to identify operational requirements, provide feedback on reusable infrastructure building blocks, and continuously improve the reliability and operability of the environment.

Skills

Site reliability
DevOps
AWS GovCloud
OpenShift ROSA
Terraform
Ansible
GitLab CI/CD
Jenkins
Python scripting

Education

Bachelor's Degree

Tools

Dynatrace
Datadog
Splunk
OpenTelemetry
AWS Solutions Architect
ROSA Certification
Red Hat ROSA
Red Hat SA OpenShift

Job description

Peraton is seeking a Site Reliability Engineer to join a remote evening-shift team responsible for AWS Commercial and GovCloud production systems, ROSA/OpenShift, and Azure/GCP knowledge. The role emphasizes operating reliability, capacity planning, and everything-as-code automation to improve deployment and recovery efficiency.

The ideal candidate has 7+ years in SRE/DevOps, must be a US citizen with Public Trust clearance, and will work with a broad observability stack to ensure top-tier

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Evening SRE: ROSA & AWS Observability Expert
Remote Evening SRE: ROSA & AWS Observability Expert

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Evening SRE - Remote Cloud Reliability Engineer (AWS/ROSA)
Evening SRE - Remote Cloud Reliability Engineer (AWS/ROSA)

Peraton • Herndon (VA)

On-site
USD 104,000 - 166,000
Night-Shift SRE (Remote) – AWS/ROSA & Observability
Night-Shift SRE (Remote) – AWS/ROSA & Observability

Peraton • Herndon (VA)

On-site
USD 104,000 - 166,000
Evening SRE – Remote Cloud & OpenShift Reliability Engineer
Evening SRE – Remote Cloud & OpenShift Reliability Engineer

Peraton • Reston (VA)

On-site
USD 104,000 - 166,000
Remote Night-Shift SRE — Cloud Infra Reliability
Remote Night-Shift SRE — Cloud Infra Reliability

Peraton • Reston (VA)

On-site
USD 104,000 - 166,000
Remote SRE - AWS/OpenShift Reliability Engineer (Evening)
Remote SRE - AWS/OpenShift Reliability Engineer (Evening)

Peraton • Northern (KY)

Hybrid
USD 120,000 - 180,000
Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer
Remote Night-Shift SRE: AWS/GovCloud Reliability Engineer

Peraton • Northern (KY)

Hybrid
USD 104,000 - 166,000
Remote Night-Shift SRE - AWS/OpenShift Reliability
Remote Night-Shift SRE - AWS/OpenShift Reliability

Peraton • United States

On-site
USD 104,000 - 166,000
Medical benefits
Dental benefits
Vision benefits
+3
Site Reliability Engineer (SRE) – Evening Shift
Site Reliability Engineer (SRE) – Evening Shift

Peraton • Northern (KY)

Hybrid
USD 120,000 - 180,000
Site Reliability Engineer (SRE) - Evening Shift
Site Reliability Engineer (SRE) - Evening Shift

Peraton • United States

On-site
USD 104,000 - 166,000