A financial institution is seeking a proactive NOC Site Reliability Engineer (SRE) to ensure the reliability and security of its IT infrastructure. The successful candidate will manage AWS cloud resources, optimize systems, and perform Root Cause Analysis (RCA). They should have at least 3 years of experience in NOC or SRE roles with hands-on expertise in cloud technologies and networking. Strong collaboration skills and a focus on compliance with internal policies are essential for this position.
Qualifications
Minimum 3 years of experience in NOC, SRE, or cloud/infrastructure operations, preferably in the financial sector.
Proven experience with AWS cloud, VMware, Kubernetes, and containerized application management.
Familiarity with regulatory requirements and IT operations compliance frameworks.
Responsibilities
Monitor, manage, and optimize IT infrastructure, including MPLS links and firewalls.
Maintain and monitor AWS cloud resources and Kubernetes clusters.
Perform Root Cause Analysis (RCA) for incidents and provide actionable recommendations.
Skills
AWS cloud
MPLS networks
Kubernetes
Monitoring tools
Incident management
Education
Bachelor’s degree in Information Technology, Computer Science, or a related field
Tools
SolarWinds NMS
Grafana
Prometheus
Elasticsearch
AWS CloudWatch
Job description
A financial institution is seeking a proactive NOC Site Reliability Engineer (SRE) to ensure the reliability and security of its IT infrastructure. The successful candidate will manage AWS cloud resources, optimize systems, and perform Root Cause Analysis (RCA). They should have at least 3 years of experience in NOC or SRE roles with hands-on expertise in cloud technologies and networking. Strong collaboration skills and a focus on compliance with internal policies are essential for this position.