Senior Site Reliability Engineer

Federal Reserve Bank of Boston

Boston (MA)

On-site

USD 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

An established industry player is seeking a Senior Engineer for the SRE/Production Operations team. In this role, you will operate and enhance production environments, ensuring reliability and scalability of services. The ideal candidate will leverage their expertise in AWS, CI/CD, and automation tools to drive continuous improvement and maintain seamless operations. This is an exciting opportunity to contribute to transformative projects within a forward-thinking organization, where your skills in cloud technologies and system architecture will be invaluable. Join a team that values innovation and collaboration, and make a significant impact in the financial services sector.

Qualifications

  • 5+ years of SRE experience in an enterprise cloud-based system.
  • Proficient in AWS, Linux, and CI/CD tools like GitLab CI.

Responsibilities

  • Operate the production environment and support ongoing technical needs.
  • Architect and implement monitoring solutions for capacity planning.

Skills

Communication Skills
AWS
Hashicorp Terraform
Ansible
Python
Linux
Docker
CI/CD
Observability Tools
Fault Injection Tooling

Education

Bachelor's degree in Computer Science

Tools

GitLab
CloudFormation
Kubernetes
Grafana
Prometheus

Job description

4 days ago Be among the first 25 applicants

All applicants must be US Citizens or Green Card holders who have resided in the US for the past 3 years. Candidates are allowed to work in one of the 12 Fed Reserve Bank locations (Boston, New York, San Francisco, Cleveland, Atlanta, Philadelphia, Dallas, Minneapolis, Richmond, Kansas City or Chicago) on a fulltime onsite basis.

Federal Reserve Financial Services (FRFS) delivers a suite of payments services to financial institutions via FedLine Solutions, FedNowSM, Fedwire, National Settlement Service (NSS), FedCash, FedACH (Automated Clearing House), and Check Services. We are currently leading a strategic effort to transform FRFS to a national, enterprise-focused organization. Through our evolved structure, we will meet the needs of the marketplace for new products and services more quickly, seek to provide a more robust and unified customer experience across our financial service.

Candidates must live within commuting distance of one of our 12 Reserve Bank locations

  • As a Senior Engineer of the SRE / Production Operations team, you will operate the production environment for the program.
  • You will help architect, implement, and leverage solution monitoring and tooling to be used for capacity planning, utilization reporting, and scaling.
  • The team uses open source and proprietary software to support Engineering, DevOps, and DevSecOps tools, services, and solutions.
  • CI/CD and IaC Pipeline automation design and development.
  • Resiliency, DR and BCP (including testing).
  • The SRE / Production Operations team is part of the Technical Operations (TechOps) department and has the overall responsibility for the design, management and execution of operations required to support the ongoing technical and delivery needs as well as the transition to production support and operations.
  • This team interfaces with internal stakeholders, customers for planning, delivery, and service management.
  • It owns ongoing ITIL processes, and the implementation and driving of continuous improvement initiatives.
  • You will work closely with Engineers and Architects in order to maintain seamless automation across the entire platform.
  • Proactively identify suspected gaps in system architecture and design experiments to expose them.
  • The ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based highly available, high performing applications.

Key Skills

  • Strong communication and collaboration skills.
  • Extensive knowledge and understanding of working in AWS environments & services.
  • Hashicorp Terraform, Consul, Vault, and Ansible.
  • Automation experience preferably using GitLab.
  • Experience with scripting languages preferably Python for automated processes.
  • Experience working in Linux environment and shell scripting.
  • Experience supporting infrastructure for large multi-services applications.
  • Experience working with continuous deployment in micro-services architectures.
  • Experience working with Docker, Containers, ECR and EKS.
  • Observability - CloudWatch, OpenSearch, Dynatrace, Grafana, Prometheus.
  • Familiarity with Fault Injection tooling (i.e. AWS Fault Injection Simulator, Gremlin, ChaosToolkit, Chaos Monkey).
  • Automation mindset to enable consistency and dependability in common actions.

Required Qualifications:

  • Bachelor's degree in Computer Science, Engineering, Mathematics, Information Technology, or related field.
  • Minimum 5 years of SRE experience in an enterprise cloud-based system.
  • Proficient with Linux/Unix systems and scripting languages (Bash, Python, etc.).
  • Proficient in AWS.
  • Proficient in Git and GitOps.
  • Experience with infrastructure-as-code tools like Terraform, and CloudFormation.
  • Strong knowledge of containerization (Docker, Kubernetes) and orchestration.
  • Expertise in CI/CD tools like GitLab CI.
  • Experience with configuration management tools like Ansible.

Additional Qualifications:

  • Excellent troubleshooting skills with the ability to quickly identify and resolve system performance and reliability issues.
  • Strong written and verbal communication skills, with the ability to explain technical concepts to non-technical stakeholders.
  • Experience in working with sensitive data.
Seniority level
  • Mid-Senior level
Employment type
  • Full-time
Job function
  • Information Technology
  • Industries
  • Data Infrastructure and Analytics
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Federal Reserve Bank of New York • Boston (MA)

On-site
USD 140,000 - 211,000
Cloud AWS Support Reliability Engineer (SRE)
Cloud AWS Support Reliability Engineer (SRE)

Federal Reserve System • New York (NY)

On-site
USD 160,000 - 230,000
Sr. SRE - Site Reliability Engineer
Sr. SRE - Site Reliability Engineer

Charles Schwab Corporation • Austin (TX)

On-site
USD 140,000 - 190,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Bank of America • Chandler (AZ)

On-site
USD 180,000 - 240,000
Cloud AWS Support Reliability Engineer (SRE)
Cloud AWS Support Reliability Engineer (SRE)

Federal Reserve Bank of New York • New York (NY)

On-site
USD 160,000 - 230,000
Educational assistance
Onsite Health & Wellness Center
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Bank of America • Charlotte (NC)

On-site
USD 180,000 - 230,000
SRE Engineer
SRE Engineer

Spatial Front, Inc • Arlington (VA)

On-site
USD 100,000 - 130,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

State of Wisconsin Investment Board • Madison (WI)

On-site
USD 150,000 - 190,000
Senior SRE, Software Engineering (AWS / Scaling Infrastructure)
Senior SRE, Software Engineering (AWS / Scaling Infrastructure)

PulseRise Technologies • New York (NY)

Hybrid
USD 130,000 - 160,000
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

VITG • Ellicott City (MD)

Hybrid
USD 90,000 - 120,000
401(k) with employer contribution
Medical/Dental/Vision insurance
Paid vacation (PTO)