SRE: Build Resilient, Self-Healing Cloud Systems

Beyond SOF

Reston (VA)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Beyond SOF is seeking a Site Reliability Engineer to join our team and work with the DoD on building robust, resilient infrastructure. You will analyze redundancy, implement monitoring, and automate tasks to reduce toil across Linux cloud-based stacks.

Leverage your experience with Docker, Kubernetes, and IaC (Puppet, Terraform, Ansible) to troubleshoot complex systems, and aim for self-healing architectures while pursuing DoD 8570 IAT Level II Certification within 30 days of hire.

Qualifications

  • 2+ years of experience working in Linux environments.
  • 1+ years of experience supporting production enterprise applications.
  • Experience with container technologies, including Docker and Kubernetes.
  • Experience with scripting, declarative Infrastructure as Code tools, including Puppet, Terraform, and Ansible.
  • Ability to dive deep into all aspects of the stack to identify and fix problems and troubleshoot.
  • Ability to obtain a DoD 8570 IAT Level II Certification, including Security+, within 30 days of hire.

Responsibilities

  • Build robust, resilient infrastructure while collaborating with the DoD.
  • Analyze redundancy, implement monitoring, and automate wherever possible.
  • Reduce toil by scripting routine tasks and enabling self-repair across the stack.

Skills

Linux
Troubleshooting
SRE fundamentals

Tools

Docker
Kubernetes
Puppet
Terraform
Ansible
Red Hat Satellite
OpenShift
AWS
Azure

Job description

Beyond SOF is seeking a Site Reliability Engineer to join our team and work with the DoD on building robust, resilient infrastructure. You will analyze redundancy, implement monitoring, and automate tasks to reduce toil across Linux cloud-based stacks.

Leverage your experience with Docker, Kubernetes, and IaC (Puppet, Terraform, Ansible) to troubleshoot complex systems, and aim for self-healing architectures while pursuing DoD 8570 IAT Level II Certification within 30 days of hire.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (TS/SCI)
Site Reliability Engineer (TS/SCI)

Beyond SOF • Reston (VA)

On-site
USD 120,000 - 160,000
DevOps Site Reliability Engineer (SRE)
DevOps Site Reliability Engineer (SRE)

IT Veterans • Washington

On-site
USD 120,000 - 180,000
Hybrid DevOps Site Reliability Engineer (TS/SCI)
Hybrid DevOps Site Reliability Engineer (TS/SCI)

Beyond SOF • Reston (VA)

Hybrid
USD 130,000 - 190,000
Remote Cloud SRE: Kubernetes, Terraform & DoD Security
Remote Cloud SRE: Kubernetes, Terraform & DoD Security

NALEJ • Arlington (VA)

On-site
USD 120,000 - 180,000
SRE: Build Scalable, Reliable Cloud Systems
SRE: Build Scalable, Reliable Cloud Systems

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Health, dental, vision insurance
Generous PTO
Learning & development
+2
SRE DevOps Engineer - 99.9% Uptime, Multi-Cloud
SRE DevOps Engineer - 99.9% Uptime, Multi-Cloud

IT Veterans • Washington

On-site
USD 120,000 - 180,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

myBridge Corporation • Austin (TX)

On-site
USD 120,000 - 160,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

CruitZi • Washington

Hybrid
USD 140,000 - 180,000
Staff SRE with TS/SCI Clearance - Cloud & Automation
Staff SRE with TS/SCI Clearance - Cloud & Automation

Okta • Virginia Beach (VA)

Hybrid
USD 174,000 - 238,000
Tier 2 SRE & Operations Specialist: Cloud Resilience
Tier 2 SRE & Operations Specialist: Cloud Resilience

Phase2 Technology • Chantilly (VA)

On-site
USD 62,000 - 141,000