Site Reliability Engineer

Talentify

Chicago (IL)

On-site

USD 110,000 - 140,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Talentify seeks a highly capable Site Reliability Engineer to join our SRE team in Chicago. You will own MFT processes, modernize legacy systems, and drive automation across enterprise infrastructure, with a focus on observability, incident management, and cloud migrations.

You'll collaborate with global technology teams, contribute to Splunk and Dynatrace integrations, and deliver reliable, scalable solutions using PowerShell, Azure CLI, and RESTful APIs in an Agile environment.

Qualifications

  • 5+ years of experience working in SRE duties.
  • Demonstrated expertise in observability, monitoring, alerting, and troubleshooting within complex technical environments.
  • Strong hands-on experience with monitoring and enterprise observability tools such as Splunk and Dynatrace
  • Solid understanding of incident management, root cause analysis, and developing effective resolution workflows.
  • Proficiency in scripting and automation for operational tasks (PowerShell, Azure CLI, or similar).
  • Microsoft Azure and AWS experience
  • Experience working within Agile or DevOps teams.
  • Knowledge of enterprise security, compliance, and governance in cloud environments

Responsibilities

  • Managed File Transfer (MFT) Ownership
  • Develop and maintain file transfer processes across 100+ internal flows for critical daily financial file transfers.
  • Modernize legacy systems, including script refactoring and secure internal vault integration.
  • Collaborate with Global Technology file transfer teams to ensure secure and reliable operations.
  • .NET Logging & Monitoring Libraries
  • Maintain and enhance common .NET libraries used for application and ETL logging.
  • Deliver automation solutions in powershell for ETL and operational tasks.
  • Integrate logging solutions across distributed systems such as Splunk, Dynatrace and Azure Log Analytics Workspace.
  • API & Data Ingestion
  • Modernize and support data ingestion pipelines using RESTful APIs and automation frameworks.
  • Cloud & Infrastructure Management
  • Contribute to Azure and AWS cloud migration efforts.
  • Support Splunk, Dynatrace and Azure Log Analytics workspace integrations.
  • Perform limited Windows servers Administration duties including software patching.
  • Operational Support Team
  • Handle certificate lifecycle management and software renewals.
  • Participate in L2/L3 incident escalations, triage, and outage management.
  • Assist in disaster recovery planning and execution.
  • Implement regular software security patch updates
  • Cross-Team Collaboration
  • Communicate effectively across infrastructure, development, and business teams.
  • Take ownership of assigned .NET or MFT projects and deliver with minimal oversight.

Skills

SRE experience
Observability
Incident management
Automation scripting
Azure & AWS
Agile/DevOps
Security/compliance knowledge

Tools

Splunk
Dynatrace
Azure CLI
PowerShell

Job description

Project Description

We are seeking a highly capable engineer to join our dynamic SRE team. This role is ideal for someone with a strong background in enterprise infrastructure, scripting, and modernization initiatives. You’ll work closely with peers and senior engineers to manage and evolve our infrastructure landscape, support cloud migrations, and independently deliver on key automation and integration projects. client offers engineers access to Claude Code and Copilot integrations to help deliver solutions efficiently.

What you can expect:
  • Managed File Transfer (MFT) Ownership
  • Develop and maintain file transfer processes across 100+ internal flows for critical daily financial file transfers.
  • Modernize legacy systems, including script refactoring and secure internal vault integration.
  • Collaborate with Global Technology file transfer teams to ensure secure and reliable operations.
  • .NET Logging & Monitoring Libraries
  • Maintain and enhance common .NET libraries used for application and ETL logging.
  • Deliver automation solutions in powershell for ETL and operational tasks.
  • Integrate logging solutions across distributed systems such as Splunk, Dynatrace and Azure Log Analytics Workspace.
  • API & Data Ingestion
  • Modernize and support data ingestion pipelines using RESTful APIs and automation frameworks.
  • Cloud & Infrastructure Management
  • Contribute to Azure and AWS cloud migration efforts.
  • Support Splunk, Dynatrace and Azure Log Analytics workspace integrations.
  • Perform limited Windows servers Administration duties including software patching.
  • Operational Support Team
  • Handle certificate lifecycle management and software renewals.
  • Participate in L2/L3 incident escalations, triage, and outage management.
  • Assist in disaster recovery planning and execution.
  • Implement regular software security patch updates
  • Cross-Team Collaboration
  • Communicate effectively across infrastructure, development, and business teams.
  • Take ownership of assigned .NET or MFT projects and deliver with minimal oversight.
What you will bring:
  • 5+ years of experience working in SRE duties.
  • Demonstrated expertise in observability, monitoring, alerting, and troubleshooting within complex technical environments.
  • Strong hands-on experience with monitoring and enterprise observability tools such as Splunk and Dynatrace
  • Solid understanding of incident management, root cause analysis, and developing effective resolution workflows.
  • Proficiency in scripting and automation for operational tasks (PowerShell, Azure CLI, or similar).
  • Microsoft Azure and AWS experience
  • Experience working within Agile or DevOps teams.
  • Knowledge of enterprise security, compliance, and governance in cloud environments
Reliability & Uptime:
  • Ensure systems are available and resilient, often measured by Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
  • Incident Response: Handle outages and performance issues, conduct postmortems, and implement fixes to prevent recurrence.
  • Automation: Replace manual operations with automated solutions (e.g., for deployments, monitoring, scaling).
  • Monitoring & Observability: Build and maintain systems to monitor application health, performance, and usage.
  • Capacity Planning & Performance Optimization: Ensure infrastructure can handle current and future loads efficiently.
  • Collaboration: Work closely with development and operations teams to improve system design and deployment processes.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender, identity, national origin, disability, or protected veteran status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

On-site
USD 100,000 - 135,000
Site Reliability Engineer
Site Reliability Engineer

Ethos Group • Irving (TX)

On-site
USD 110,000 - 160,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mikealbert • Cincinnati (OH)

On-site
USD 100,000 - 130,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink • United States

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GovCIO • Arlington (VA)

On-site
USD 210,000 - 230,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • United States

On-site
USD 140,000 - 190,000
Sr SRE Automation Engineer
Sr SRE Automation Engineer

Compunnel, Inc. • Austin (TX), Northern (KY)

On-site
USD 130,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Pacificacontinental • Pacifica (CA)

On-site
USD 140,000 - 190,000
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)

Mindlance • Irving (TX)

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

IntraEdge • Austin (TX)

On-site
USD 120,000 - 180,000