Senior DevOps Engineer / Site Reliability Engineer (SRE)

KellyMitchell Group

Chicago (IL)

On-site

USD 87,000 - 124,000

Full time

6 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical, Dental, & Vision Insurance
Employee-Owned Profit Sharing (ESOP)
401K offered

Job summary

KellyMitchell Group seeks a Senior DevOps Engineer / Site Reliability Engineer to support enterprise backup, cyber recovery, and platform resiliency initiatives in a hybrid Chicago setting. This role emphasizes automation, SRE practices, and scalable recovery architectures across multi-cloud and on-prem environments.

The ideal candidate has 7+ years in backup/infrastructure/SRE, expertise in major backup tools, cloud platforms, IaC, and observability tooling, and a strong security-first mindset

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.
  • 7+ years of experience in Backup Engineering, Infrastructure Engineering, Site Reliability Engineering, or a related field.
  • 5+ years of experience designing and supporting enterprise backup solutions.
  • 3+ years of experience supporting cyber recovery architectures.
  • Strong understanding of distributed systems and highly available infrastructure.
  • Experience implementing SRE methodologies within enterprise environments.
  • Experience with backup technologies including Cohesity, Rubrik, Dell PowerProtect Data Manager, Dell Data Domain, Dell Cyber Recovery, Commvault, Veritas NetBackup, or Veeam.
  • Experience with cyber recovery solutions including air-gapped vaults, immutable backups, clean rooms, IRE, recovery orchestration, and ransomware recovery.
  • Experience with Microsoft Azure, AWS, Google Cloud Platform, VMware, Kubernetes, OpenShift, Linux, Windows Server, Active Directory, and enterprise storage platforms.
  • Proficiency with Ansible, Terraform, Python, PowerShell, Bash, GitHub, GitHub Actions, and CI/CD pipelines.
  • Experience with monitoring tools such as Dynatrace, Grafana, Prometheus, Splunk, ELK Stack, ServiceNow.

Responsibilities

  • Engineer and support highly available enterprise platforms using SRE best practices.
  • Define and measure SLIs, SLOs, and error budgets for backup and recovery services.
  • Develop automation to reduce operational overhead and improve reliability.
  • Conduct root cause analysis and implement permanent corrective actions.
  • Improve system scalability, performance, resiliency, and recoverability.
  • Support incident response and major recovery efforts.
  • Design and administer enterprise backup and recovery solutions.
  • Engineer immutable backup architectures for ransomware resilience.
  • Develop backup strategies across virtual, physical, cloud-native, and enterprise environments.
  • Optimize backup performance, retention, replication, encryption, and recovery objectives.
  • Implement policy-based automation and lifecycle management.
  • Ensure compliance with RPO and RTO requirements.
  • Design and implement cyber recovery solutions including air-gapped vaults and IRE.
  • Collaborate with cybersecurity teams to strengthen recovery readiness.
  • Develop IaC and Recovery as Code solutions.
  • Build automated recovery workflows using Ansible, Terraform, Python, PowerShell, Bash, GitHub Actions.
  • Implement enterprise monitoring and observability solutions.
  • Create dashboards and reporting for stakeholders.

Skills

Ansible
Terraform
Python
PowerShell
Bash
GitHub Actions
SRE practices
Backup technologies
Azure
AWS
GCP
Kubernetes
OpenShift
Linux
Windows Server
Active Directory

Education

Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience

Tools

Dynatrace
Grafana
Prometheus
Splunk
ELK Stack
ServiceNow

Job description

Job Summary

Our client, a leader financial services provider, is seeking a Senior DevOps Engineer / Site Reliability Engineer (SRE) to support enterprise backup, cyber recovery, and platform resiliency initiatives. The ideal candidate brings experience in backup and recovery solutions, cyber resiliency, infrastructure automation, and SRE practices within large-scale enterprise environments.

Contract Details

Contract-to-hire potential Location: Chicago, IL preferred, and Tempe, AZ would be considered Work model: Hybrid

Core Responsibilities
  • Engineer and support highly available, resilient enterprise platforms utilizing SRE best practices
  • Define and measure SLIs, SLOs, and error budgets for backup and recovery services
  • Develop automation to reduce operational overhead and improve reliability
  • Conduct root cause analysis and implement permanent corrective actions
  • Improve system scalability, performance, resiliency, and recoverability
  • Support incident response and major recovery efforts
  • Design, implement, and administer enterprise backup and recovery solutions
  • Engineer immutable backup architectures to support ransomware resilience
  • Develop backup strategies across virtual, physical, database, cloud-native, and enterprise application environments
  • Optimize backup performance, retention, replication, encryption, and recovery objectives
  • Implement policy-based automation and lifecycle management
  • Ensure compliance with RPO and RTO requirements
  • Design and implement cyber recovery solutions, including air-gapped recovery vaults, clean room environments, Isolated Recovery Environments (IRE), and immutable storage architectures
  • Support ransomware resiliency and cyber recovery initiatives
  • Develop secure recovery workflows and automated validation processes
  • Collaborate with cybersecurity teams to strengthen recovery readiness
  • Develop Infrastructure as Code (IaC) and Recovery as Code solutions
  • Build automated recovery workflows utilizing Ansible, Terraform, Python, PowerShell, and GitHub Actions
  • Implement enterprise monitoring and observability solutions
  • Create dashboards and reporting for operational and executive stakeholders
Required Skills & Experience (Must-Haves)
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience
  • 7+ years of experience in Backup Engineering, Infrastructure Engineering, Site Reliability Engineering, or a related field
  • 5+ years of experience designing and supporting enterprise backup solutions
  • 3+ years of experience supporting cyber recovery architectures
  • Strong understanding of distributed systems and highly available infrastructure
  • Experience implementing SRE methodologies within enterprise environments
  • Experience with one or more backup technologies, including Cohesity, Rubrik, Dell PowerProtect Data Manager, Dell Data Domain, Dell Cyber Recovery, Commvault, Veritas NetBackup, or Veeam
  • Experience with cyber recovery solutions, including air-gapped vaults, immutable backups, clean rooms, Isolated Recovery Environments (IRE), recovery orchestration, recovery validation, cyber resiliency testing, and ransomware recovery
  • Experience with Microsoft Azure, AWS, Google Cloud Platform, VMware, Kubernetes, OpenShift, Linux, Windows Server, Active Directory, and enterprise storage platforms
  • Proficiency with Ansible, Terraform, Python, PowerShell, Bash, GitHub, GitHub Actions, and CI/CD pipelines
  • Experience with Dynatrace, Grafana, Prometheus, Splunk, ELK Stack, ServiceNow, or similar monitoring and observability tools
  • Knowledge of Zero Trust Architecture, NIST Cybersecurity Framework, CIS Controls, IAM, MFA, and encryption and key management practices
Preferred Skills & Experience (Nice-to-Haves)
  • Experience within financial services or other highly regulated industries
  • Experience supporting GSIB cyber resiliency programs
  • Familiarity with Federal Reserve, OCC, or FFIEC regulatory requirements
  • Experience with chaos engineering and resilience testing
  • Experience with AIOps and predictive analytics
  • Experience leveraging enterprise reliability metrics and SRE tooling
Key Competencies & Behaviors
  • Systems Thinking
  • Reliability Mindset
  • Analytical Problem Solving
  • Cross-Functional Leadership
  • Technical Communication
  • Recovery Planning
  • Automation Focus
  • Continuous Improvement
Work Environment

Location: Chicago, IL preferred, and Tempe, AZ would be considered Work model: Hybrid

Compensation & Benefits

Pay Range: The approximate pay range for this position is between $63.00 and $90.00. Please note that the pay range provided is a good faith estimate. Final compensation may vary based on factors including but not limited to background, knowledge, skills, and location. We comply with local wage minimums.

  • Medical, Dental, & Vision Insurance Plans
  • Employee-Owned Profit Sharing (ESOP)
  • 401K offered
About KellyMitchell

At KellyMitchell, our culture is world class. We’re movers and shakers! We don’t mind a bit of friendly competition, and we reward hard work with unlimited potential for growth. This is an exciting opportunity to join a company known for innovative solutions and unsurpassed customer service. We're passionate about helping companies solve their biggest IT staffing & project solutions challenges. As an employee-owned, women-led organization serving Fortune 500 companies nationwide, we deliver expert service at a moment's notice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Enterprise Backup & Cyber Recovery
Senior SRE: Enterprise Backup & Cyber Recovery

KellyMitchell Group • Chicago (IL)

Hybrid
USD 87,000 - 124,000
Medical, Dental, & Vision Insurance
Employee-Owned Profit Sharing (ESOP)
401K offered
Information Technology - DevOps Engineer (Cloud Engineer)
Information Technology - DevOps Engineer (Cloud Engineer)

Accord Technologies Inc • Chicago (IL)

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineer / Backup Engineer
Senior Site Reliability Engineer / Backup Engineer

ICON Technologies • Chicago (IL)

Hybrid
USD 120,000 - 160,000
SRE/Backup Engineer
SRE/Backup Engineer

Spectraforce Technologies • Chicago (IL)

On-site
USD 120,000 - 180,000
Sr. Backup Engineer - Cyber Resiliency
Sr. Backup Engineer - Cyber Resiliency

The Judge Group • Chicago (IL)

Hybrid
USD 120,000 - 160,000
Backup/Data Protection Infrastructure Analyst
Backup/Data Protection Infrastructure Analyst

Synergis • Atlanta (GA)

Hybrid
USD 90,000 - 135,000
Pension
401K match
Medical
+1
Cyber Recovery Engineer / Disaster Recovery Engineer
Cyber Recovery Engineer / Disaster Recovery Engineer

Prairie Consulting Services • Illinois

Hybrid
USD 119,000 - 170,000
Senior SRE - Cyber Recovery & Backup Engineer
Senior SRE - Cyber Recovery & Backup Engineer

Spectraforce Technologies • Chicago (IL)

On-site
USD 120,000 - 180,000
Senior SRE & Cyber Recovery Engineer (Hybrid)
Senior SRE & Cyber Recovery Engineer (Hybrid)

ICON Technologies • Chicago (IL)

Hybrid
USD 120,000 - 160,000
Backup Engineer
Backup Engineer

Prairie Consulting Services • Chicago (IL)

Hybrid
USD 120,000 - 160,000