Site Reliability Engineer

oneazcreditunion

Phoenix (AZ)

On-site

USD 110,000 - 140,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OneAZ Credit Union is seeking a Site Reliability Engineer in Phoenix, AZ to ensure reliability, availability, and performance of technology platforms. The role blends infrastructure engineering, automation, monitoring, and resiliency for on‑premises and cloud environments.

You will design automation scripts (PowerShell, Python, Bash), manage SolarWinds, and lead disaster recovery testing and recovery planning with cross‑functional teams.

Qualifications

  • High school diploma required.
  • Bachelor’s degree or equivalent experience in a related technical field.
  • 5–8 years of experience supporting enterprise infrastructure environments and monitoring platforms.
  • Strong knowledge of networking, cloud, Windows Server, virtualization, and storage.

Responsibilities

  • Administer and optimize SolarWinds and monitoring platforms.
  • Develop and maintain monitoring dashboards, alerts, and performance metrics.
  • Lead disaster recovery planning, testing, and recovery validation activities.
  • Coordinate failover/failback for critical apps and infrastructure.
  • Partner with teams to improve resilience and incident response.

Skills

Networking
Cloud technologies
Windows Server
Virtualization
Storage
Disaster recovery
PowerShell
Python/Bash

Education

Bachelor's Degree in Information Technology, Computer Science, Information Systems, Engineering, or related field; or equivalent

Tools

SolarWinds
Azure

Job description

Join Us in Making an Impact

At OneAZ Credit Union, our success is measured only by yours. We’re here to create lasting change in the lives of our members, our communities, and our team. If you’re looking for a career with purpose, where your work truly matters-you’ve found it!

Who You Are

You’re impactful, compassionate, and fearless, ready to embrace new challenges and shape the future of financial well-being. You take accountability for our success and thrive in an environment where curiosity is celebrated. If this sounds like you, let’s build something great together.

What You’ll Do

This position will be located at our Corporate Office: 2355 W Pinnacle Peak Rd, Phoenix, AZ 85027

The Site Reliability Engineer is responsible for ensuring the reliability, availability, performance, and recoverability of OneAZ’s technology platforms and infrastructure. This position combines infrastructure engineering, automation, monitoring, and resiliency practices to maintain highly available systems supporting associates and members. The engineer designs and implements automation solutions using PowerShell and other scripting technologies, administers enterprise monitoring platforms, leads disaster recovery testing activities, and partners with technology teams to improve operational resilience, service reliability, and recovery readiness across on-premises and cloud environments. This role serves as a key contributor to incident response, infrastructure modernization, and continuous improvement initiatives focused on reducing operational risk and improving system uptime.

The Site Reliability Engineer works closely with Infrastructure, Information Security, Application Support, Enterprise Architecture, and business teams to identify operational risks, strengthen recovery capabilities, and improve the overall resilience of technology services that support associates and members.

Essential Functions
  • Administer, maintain, and optimize SolarWinds and other enterprise monitoring platforms.
  • Develop and maintain monitoring dashboards, alerts, reports, and performance metrics.
  • Proactively monitor and improve infrastructure health, availability, capacity, and performance, identifying and mitigating potential issues before they impact service reliability or user experience.
  • Proactively identify infrastructure risks, reliability concerns, and opportunities for improvement.
  • Lead disaster recovery planning, testing, failover exercises, and recovery validation activities.
  • Develop and maintain disaster recovery documentation, runbooks, recovery procedures, and test plans.
  • Coordinate and execute system failover and failback activities for critical applications and infrastructure.
  • Partner with Infrastructure, Security, and Application teams to improve system resiliency and operational readiness.
  • Support business continuity planning initiatives and technology recovery efforts.
  • Lead the investigation and resolution of complex infrastructure and service reliability incidents, conduct root cause analyses (RCAs), identify underlying systemic issues, and drive corrective and preventive actions to improve availability, performance, resiliency and operational excellence.
  • Track, document, and report on disaster recovery testing results, remediation activities, and recovery readiness metrics.
  • Maintain infrastructure diagrams, recovery documentation, and operational procedures.
  • Support audits, examinations, and compliance activities related to disaster recovery, resiliency, and infrastructure operations.
  • Research and recommend technologies and best practices that improve reliability, monitoring, and recoverability.
  • Participate in infrastructure maintenance activities, upgrades, and projects.
  • Develop and maintain automation scripts using PowerShell, Python, Bash, or similar tools to streamline infrastructure operations, automate remediation of common issues, and enhance service availability and performance.
What You Bring
  • High School Diploma Required
  • Bachelor's Degree in Information Technology, Computer Science, Information Systems, Engineering, or a related technical field; or equivalent combination of education and experience. Required
  • 5-8 years similar or related experience of experience supporting enterprise infrastructure environments and administering infrastructure monitoring platforms. Required
  • Deep knowledge of networking and cloud technologies, Windows Server, virtualization, and storage with hands on experience supporting and optimizing enterprise infrastructure environments.
  • Experience with disaster recovery testing, failover planning, or business continuity initiatives.
  • Experience administering SolarWinds in a large enterprise environment. Preferred
  • Experience within a financial institution, credit union, or regulated industry. Preferred
  • Experience supporting hybrid cloud environments including Microsoft Azure. Preferred
  • Experience coordinating disaster recovery exercises and recovery read
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

OneAZ Credit Union • Phoenix (AZ)

On-site
USD 97,000 - 121,000
Generous paid time off
Salary range disclosed
Site Reliability Engineer - Reliability, DR & Automation
Site Reliability Engineer - Reliability, DR & Automation

oneazcreditunion • Phoenix (AZ)

On-site
USD 110,000 - 140,000
Site Reliability Engineer – Cloud, DR & Automation
Site Reliability Engineer – Cloud, DR & Automation

OneAZ Credit Union • Phoenix (AZ)

On-site
USD 97,000 - 121,000
Generous paid time off
Salary range disclosed
Network Engineer
Network Engineer

oneazcreditunion • Phoenix (AZ)

On-site
USD 110,000 - 150,000
Network Engineer
Network Engineer

OneAZ Credit Union • Phoenix (AZ)

On-site
USD 97,000 - 121,000
Generous paid time off
Senior Information Security Analyst
Senior Information Security Analyst

OneAZ Credit Union • Phoenix (AZ)

On-site
USD 84,000 - 105,000
PTO
Childcare assistance
401K
+3
Database Administrator
Database Administrator

oneazcreditunion • Phoenix (AZ)

On-site
USD 90,000 - 130,000
Senior Systems Analyst
Senior Systems Analyst

OneAZ Credit Union • Glendale (AZ)

On-site
USD 97,000 - 121,000
Generous paid time off
Competitive compensation
On-site at Corporate Office, Phoenix
Infrastructure Administrator - Lake Havasu, AZ
Infrastructure Administrator - Lake Havasu, AZ

Arizona Federal Credit Union • Lake Havasu City (AZ)

Hybrid
USD 75,000 - 110,000
Infrastructure Administrator - Lake Havasu, AZ
Infrastructure Administrator - Lake Havasu, AZ

Arizona Financial Credit Union • Lake Havasu City (AZ)

Hybrid
USD 80,000 - 110,000
Hybrid work environment