Site Reliability Engineer (SRE) – II

Huntington Bank

Easton (PA)

On-site

USD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A financial institution in Easton, PA is seeking a Site Reliability Engineer (SRE) Level II to ensure the reliability, scalability, and performance of critical systems. You will lead incident management, develop automation solutions, and collaborate with teams on system design. The ideal candidate has a Bachelor’s degree, 3+ years of SRE or DevOps experience, and strong skills in cloud platforms like AWS and monitoring tools. The position offers in-office work with flexible arrangements.

Qualifications

  • 3+ years of experience in site reliability engineering, DevOps, or related roles.
  • Experience troubleshooting production issues.
  • Ability to balance development and support tasks.

Responsibilities

  • Maintain availability and scalability of infrastructure.
  • Lead incident management and troubleshooting efforts.
  • Develop automation scripts using Terraform and Ansible.
  • Collaborate on the design of new services and applications.
  • Build and maintain monitoring and alerting solutions.

Skills

Linux/Unix administration
Scripting (Python, Bash, Go)
Cloud platforms (AWS, GCP, Azure)
Docker and Kubernetes
Monitoring tools (Dynatrace, Prometheus, Grafana)
Networking fundamentals
CI/CD tools (Jenkins, GitLab CI)
Problem-solving skills
Customer service focus

Education

Bachelor’s degree in computer science or Information Technology

Tools

Terraform
Ansible
Kubernetes

Job description

# **Description******This employer will not sponsor applicants for the following work visas: F-1 student, H-1B worker, O-1 worker, TN worker, E-3 worker. Applicants must be currently authorized to work in the United States on a full-time basis.******Summary:**As a Site Reliability Engineer (SRE) Level II, you will play a key role in maintaining the availability, scalability, and performance of critical infrastructure and services. You will be responsible for building and automating solutions that enhance system reliability and support continuous delivery. In this role, you will handle more complex operational tasks and incidents, provide mentorship to junior SREs, and collaborate with development teams to ensure systems are designed for reliability from the ground up.* **Incident Management :*** complex incidents, and ensure service uptime.* Lead troubleshooting efforts for high-impact production issues, providing detailed root cause analysis (RCA) and preventative measures.* Participate in on-call rotations, acting as an escalation point for Level 1 SREs during major incidents.* **Automation & Infrastructure as Code (IaC):*** Develop and maintain automation scripts and infrastructure using tools like Terraform, Ansible, or CloudFormation.* Implement automation solutions to eliminate manual tasks and improve system reliability, scalability, and performance.* **Performance & Scalability:*** Analyze system performance and recommend optimizations for scalability and reliability.* Support capacity planning efforts by monitoring system metrics, traffic* patterns, and usage trends to predict future resource needs.* **System Design & Architecture:*** Collaborate with software engineering teams to influence the design of new services and applications, ensuring they are scalable, reliable, and resilient from the start.* Contribute to architectural decisions, ensuring alignment with best practices in fault tolerance, redundancy, and recovery.* **Monitoring & Observability:*** Build and maintain robust monitoring, alerting, and observability solutions to proactively detect and resolve issues before they impact end users.* Optimize existing monitoring tools (e.g., Prometheus, Grafana, Datadog, Dynatrace) and build custom dashboards for better visibility into system health.* **Security & Compliance:*** Ensure systems and infrastructure are secure, compliant, and aligned with organizational policies and industry best practices.* Assist with vulnerability management, system patching, and implementing security measures to protect the integrity and availability of services.* **Continuous Improvement:*** Lead efforts to continuously improve operational processes, tools, and workflows.* Implement and enforce best practices in deployment, monitoring, and incident management to improve overall system reliability and reduce downtime.**Basic Qualifications:*** Bachelor’s degree in computer science, Information Technology* 3+ years of experience in site reliability engineering, DevOps, systems administration, or related roles.**Preferred Qualifications:*** Strong experience with Linux/Unix administration and proficiency in scripting (e.g., Python, Bash, Go).* Deep understanding of cloud platforms (AWS, GCP, Azure) and related services (EC2, S3, Lambda, Kubernetes, etc.).* Experience with containerization and orchestration technologies like Docker and Kubernetes.* Proven track record of managing complex infrastructure, troubleshooting production issues, and optimizing system performance* Proficiency with monitoring and observability tools such as dynatrace, Prometheus, Grafana, Datadog, ELK Stack, or similar platforms.* Strong understanding of networking fundamentals (DNS, HTTP, TCP/IP), load balancing, and CDNs.* Experience with CI/CD tools (Jenkins, GitLab CI, CircleCI) and infrastructure automation (Terraform, Ansible, Puppet).* Familiarity with distributed systems and microservices architecture.* Excellent problem-solving and troubleshooting skills, especially in diagnosing production issues in high-scale environments.* Microsoft Office experience* Experience working in multi-platform environment* Ability to balance both development and support roles* Experience in working on projects that involve business segments* Strong analytical, strong troubleshooting skills and excellent communication skills* Strong interpersonal skills, focus on customer service, and the ability to work well with other IT, vendor, and business groups**Exempt Status: (Yes** = not eligible for overtime pay) (**No** = eligible for overtime pay)Yes**Workplace Type:**OfficeOur Approach to **Office** Workplace TypeCertain positions outside our branch network may be eligible for a flexible work arrangement. We’re combining the best of both worlds: in-office and work from home. Our approach enables our teams to deepen connections, maintain a strong community, and do their best work. Remote roles will also have the opportunity to come together in our offices for moments that matter. Specific work arrangements will be provided by the hiring team.Huntington will not sponsor applicants for this position for immigration benefits, including but not limited to assisting with obtaining work permission for F-1 students, H-1B professionals, O-1 workers, TN workers, E-3 workers, among other immigration statuses. Applicants must be currently authorized to work in the United States on a full-time basis.Huntington is an Equal Opportunity Employer.Tobacco-Free Hiring Practice: Visit Huntington's Career Web Site for more details.**Note to Agency Recruiters:** Huntington will not pay a fee for any placement resulting from the receipt of an unsolicited resume. All unsolicited resumes sent to any Huntington colleagues, directly or indirectly, will be considered Huntington property. Recruiting agencies must have a valid, written and fully executed Master Service Agreement and Statement of Work for consideration.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Huntington Bank • Columbus (OH)

Hybrid
USD 57,000 - 113,000
Health insurance
Wellness program
Life and disability insurance
+2
Site Reliability Engineer (SRE) – II
Site Reliability Engineer (SRE) – II

Huntington National Bank • Columbus (OH)

Hybrid
USD 90,000 - 120,000
Digital - Principal SRE
Digital - Principal SRE

Huntington Bank • Columbus (OH)

On-site
USD 120,000 - 160,000
Digital - Principal SRE (AI Engineer)
Digital - Principal SRE (AI Engineer)

Huntington • Columbus (OH)

Hybrid
USD 140,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Huntington Bancshares, Inc. • Hagerstown (MD)

On-site
USD 57,000 - 113,000
Health insurance
Wellness program
Life and disability insurance
+3
Site Reliability Engineer
Site Reliability Engineer

Huntington National Bank • Minnetonka (MN)

On-site
USD 57,000 - 113,000
Health insurance
Wellness program
Retirement savings plan
+3
Senior AI Reliability Engineer — Remote
Senior AI Reliability Engineer — Remote

Huntington Bancshares, Inc. • Columbus (OH)

Hybrid
USD 120,000 - 150,000
Lead SRE and Vulnerability Engineer
Lead SRE and Vulnerability Engineer

Huntington Bank • Columbus (OH)

Hybrid
USD 95,000 - 115,000
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer

Huntington National Bank • Columbus (OH)

On-site
USD 85,000 - 115,000
Flexible work arrangements
Collaborative work environment
Professional development opportunities
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer

Huntington National Bank • Columbus (OH)

Hybrid
USD 90,000 - 120,000