Site Reliability Engineer II

Jobtailor

San Francisco (CA)

On-site

USD 120,000 - 180,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor is seeking an experienced SRE/DevOps engineer to monitor and support AWS Commercial/GovCloud and EKS environments. You will perform real-time monitoring, triage incidents, and drive remediation of vulnerabilities.

The role emphasizes Linux administration, Bash/Python scripting, and CI/CD deployments with GitLab CI/CD and Argo tools, with US-only residency and readiness for shift rotations. Applicants should meet FedRAMP suitability requirements and be willing to work on US soil,

Qualifications

  • 3+ years in SRE, DevOps, systems administration, or cloud support.
  • Solid Linux administration with RHEL, Amazon Linux, and Ubuntu.
  • Bash/Python scripting.
  • Working knowledge of Kubernetes/EKS and AWS core services.
  • L4/L7 load balancing troubleshooting.
  • Experience with Prometheus, Grafana, and ChatOps via Slack.
  • US citizenship and US-soil residency.
  • Genuine availability for overnight/weekend rotational shifts.
  • FedRAMP suitability and US-soil-only eligibility.

Responsibilities

  • Perform real-time monitoring and first-response triage across AWS Commercial/GovCloud and EKS infrastructure
  • Conduct log-level root cause investigations in Kibana and Elasticsearch
  • Remediate CVEs and perform STIG hardening within FedRAMP SLA windows
  • Execute and troubleshoot deployments through GitLab CI/CD, ArgoCD, and Argo Workflows

Skills

Kubernetes/EKS Administration
Real-Time Monitoring
Log-Level Investigations
Vulnerability Remediation
STIG Hardening
Load Balancing Troubleshooting
Prometheus
Grafana
Bash Scripting
Python Scripting
Availability for Rotational Shifts

Tools

Kibana
Elasticsearch
GitLab CI/CD
Argo Workflows
Slack

Job description


  • Perform real-time monitoring and first-response triage across AWS Commercial/GovCloud and EKS infrastructure

  • Conduct log-level root cause investigations in Kibana and Elasticsearch

  • Remediate CVEs and perform STIG hardening within FedRAMP SLA windows

  • Execute and troubleshoot deployments through GitLab CI/CD, ArgoCD, and Argo Workflows


Requirements



  • 3+ years in SRE, DevOps, systems administration, or cloud support

  • Solid Linux administration with RHEL, Amazon Linux, and Ubuntu

  • Bash/Python scripting

  • Working knowledge of Kubernetes/EKS and AWS core services

  • L4/L7 load balancing troubleshooting

  • Experience with Prometheus, Grafana, and ChatOps via Slack

  • US citizenship and US-soil residency

  • Genuine availability for overnight/weekend rotational shifts

  • FedRAMP suitability and US-soil-only eligibility


Core Competencies


Demonstrates expertise in real-time monitoring, log-level investigations, and remediation of vulnerabilities within AWS and EKS environments. Proficient in Linux administration and scripting, with a strong focus on CI/CD processes and compliance with FedRAMP standards.


Highest-signal resume keywords



  • AWS Infrastructure Management

  • Kubernetes/EKS Administration

  • Linux Administration (RHEL, Amazon Linux, Ubuntu)

  • Bash/Python Scripting

  • CI/CD Deployment (GitLab, ArgoCD)


ATS Optimization Keywords


Hard Skills



  • Real-Time Monitoring

  • Log-Level Investigations

  • Vulnerability Remediation

  • STIG Hardening

  • Load Balancing Troubleshooting

  • Prometheus

  • Grafana

  • ChatOps

  • Bash Scripting

  • Python Scripting


Soft Skills



  • Availability for Rotational Shifts


Certifications & Qualifications



  • FedRAMP Suitability


Industry Keywords



  • SRE

  • DevOps

  • Cloud Support

  • Systems Administration

  • US Citizenship

  • US-Soil Residency


Tools & Technologies



  • Kibana

  • Elasticsearch

  • GitLab CI/CD

  • Argo Workflows

  • Slack

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

Jobtailor • Arlington (VA)

On-site
USD 120,000 - 180,000
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Jobtailor • Washington

On-site
USD 140,000 - 185,000
Senior Site Reliability Engineer – AWS/Datacenter
Senior Site Reliability Engineer – AWS/Datacenter

Jobtailor • Bellevue (WA)

On-site
USD 160,000 - 220,000
DevOps Engineer
DevOps Engineer

Jobtailor • Fort Meade (MD)

On-site
USD 120,000 - 180,000
Product Owner – DevOps
Product Owner – DevOps

Jobtailor • Missouri

On-site
USD 140,000 - 190,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobtailor • Arizona

On-site
USD 120,000 - 190,000
Staff DevOps Engineer
Staff DevOps Engineer

Jobtailor • Mountain View (CA)

Hybrid
USD 180,000 - 240,000
Cloud Virtual Infrastructure System Administrator
Cloud Virtual Infrastructure System Administrator

Jobtailor • Beavercreek (OH)

On-site
USD 90,000 - 120,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobtailor • California (MO)

On-site
USD 120,000 - 170,000
#II Site Reliability Engineer II
#II Site Reliability Engineer II

C-Serv • Reston (VA)

Hybrid
USD 96,000 - 124,000