Get more replies from employers
Send a job-specific resume in minutes.
Blue Spire Inc. in Phoenix, AZ is seeking a Senior Site Reliability Engineer to design, implement, and operate highly available platforms for critical financial services workloads.
You will optimize cloud-native stacks on AWS/Azure/GCP, container orchestration with Kubernetes/OpenShift, and infrastructure as code with Terraform, Ansible, and CloudFormation. Strong emphasis on incident management, observability, and automation to achieve operational resilience.
Extensive experience in Site Reliability Engineering (SRE) and Production Engineering
Strong expertise in AWS/Azure/GCP cloud platforms
Hands-on experience with Kubernetes, Docker, OpenShift, and container orchestration
Proficiency in Infrastructure as Code (Terraform, Ansible, CloudFormation)
Strong knowledge of CI/CD pipelines using Jenkins, GitHub Actions, Azure DevOps, or GitLab CI
Experience with Observability & Monitoring tools such as Splunk, Datadog, Dynatrace, Prometheus, Grafana, ELK, or New Relic
Expertise in Incident Management, Problem Management, Root Cause Analysis (RCA), and Service Availability
Strong scripting/programming skills in Python, Bash, or Go
Experience implementing automation, self-healing systems, and operational resilience best practices
Excellent communication and stakeholder management skills
Banking or Financial Services experience
Experience designing enterprise-scale, highly available, and fault-tolerant platforms
Exposure to DevSecOps, AI Ops, or Chaos Engineering is a plus
If you Know someone who fits this role? Feel free to like, comment, or share this opportunity with your network!