Senior Site Reliability Engineer (AWS)

Broadridge Financial Solutions

Philippines

Hybrid

PHP 1,800,000 - 2,400,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Broadridge Financial Solutions is seeking a Senior Site Reliability Engineer (AWS) to design, build, and operate highly reliable, scalable, and secure platforms across hybrid on-prem and cloud environments. The role emphasizes automation, resiliency, observability, and cost efficiency while reducing operational toil through collaboration with cross-functional teams.

You will define SLOs/SLIs, implement self-healing strategies, and drive cloud migrations with a focus on security and compliance in

Qualifications

  • 8+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Systems Engineering.
  • Strong programming experience in Python, Java, or similar languages.
  • Deep experience with Linux/Unix systems.
  • Hands-on expertise with AWS and cloud-native architectures.
  • Proven experience with Terraform and Infrastructure as Code.
  • Strong understanding of networking, security, and distributed systems.
  • Experience operating mission-critical, high-volume platforms.
  • Nice to Have Experience in financial services or highly regulated environments.
  • Experience with EKS/Kubernetes at scale.
  • Familiarity with Chaos Engineering and resilience testing.
  • Experience leading cloud cost optimization (FinOps) initiatives.
  • Prior experience transitioning traditional infrastructure teams into SRE practices.

Responsibilities

  • Design and implement high-availability, fault-tolerant architectures across on-prem and cloud platforms (AWS).
  • Lead multi-region DR planning, implementation, and testing, including RTO/RPO definition and validation.
  • Define and enforce SLOs, SLIs, and error budgets to balance reliability with delivery velocity.
  • Drive self-healing automation and proactive remediation strategies.
  • Build and maintain infrastructure using Terraform and configuration management tools (e.g., Chef).
  • Develop automation to eliminate manual operational tasks and create reusable modules and guardrails.

Skills

Python/Java
Linux/Unix
Distributed systems
SRE/DevOps
Networking & security
Chaos engineering
FinOps

Tools

Terraform
Chef
Kubernetes (EKS)
AWS services

Job description

At Broadridge, we've built a culture where the highest goal is to empower others to accomplish more. If you’re passionate about developing your career, while helping others along the way, come join the Broadridge team.

Role Overview

We are seeking a Senior Site Reliability Engineer (AWS) to design, build, and operate highly reliable, scalable, and secure platforms supporting business-critical applications across hybrid (on-prem and cloud) environments. This role blends software engineering, systems engineering, and operational excellence, with a strong focus on automation, resiliency, observability, and cost efficiency. The SRE will partner closely with application development, infrastructure, security, and product teams to reduce operational toil, improve system reliability, and enable faster, safer delivery of services.

Responsibilities
  • Reliability & Resiliency Engineering Design and implement high-availability, fault-tolerant architectures across on-prem and cloud platforms (AWS). Lead multi-region DR planning, implementation, and testing, including RTO/RPO definition and validation. Define and enforce SLOs, SLIs, and error budgets to balance reliability with delivery velocity. Drive self-healing automation and proactive remediation strategies.
  • Automation & Infrastructure as Code Build and maintain infrastructure using Terraform and configuration management tools (e.g., Chef). Develop automation to eliminate manual operational tasks (TOIL reduction). Create reusable modules, pipelines, and guardrails for standardized deployments. Automate certificate lifecycle management, key rotation, and security updates.
  • Observability & Monitoring Design and implement end-to-end observability (metrics, logs, traces, synthetic monitoring). Build dashboards, alerts, and runbooks to enable fast detection and resolution of incidents. Improve signal-to-noise ratio in alerting to reduce operational fatigue. Perform root cause analysis (RCA) and lead post-incident reviews with actionable follow-ups.
  • Cloud & Platform Engineering Engineer and operate platforms on AWS, including services such as: EKS, EC2, RDS/Aurora, Lambda, API Gateway, CloudFront, WAF, ALB/NLB, CloudWatch, X-Ray, IAM, Secrets Manager. Lead cloud migrations and modernization initiatives, including legacy system refactoring. Implement secure networking patterns (VPCs, private subnets, controlled egress).
  • Performance, Scalability & Cost Optimization Identify and resolve performance bottlenecks through testing and analysis. Drive FinOps initiatives to optimize infrastructure cost without compromising reliability. Implement capacity planning and autoscaling strategies.
  • CI/CD & SDLC Enablement Design and support CI/CD pipelines enabling safe, repeatable deployments. Embed reliability practices into the SDLC (testing, rollout strategies, rollback). Partner with development teams to improve operability of applications before production.
  • Security & Compliance Partner with security and legal teams to meet regulatory and compliance requirements (e.g., data residency, GDPR-related controls). Implement secure access controls, secrets management, and encryption best practices. Participate in security reviews, audits, and risk assessments.
Your Profile
  • 8+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, or Systems Engineering
  • Strong programming experience in Python, Java, or similar languages
  • Deep experience with Linux/Unix systems
  • Hands-on expertise with AWS and cloud-native architectures
  • Proven experience with Terraform and Infrastructure as Code
  • Strong understanding of networking, security, and distributed systems
  • Experience operating mission- critical, high-volume platforms
  • Nice to Have Experience in financial services or highly regulated environments
  • Experience with EKS/Kubernetes at scale
  • Familiarity with Chaos Engineering and resilience testing
  • Experience leading cloud cost optimization (FinOps) initiatives
  • Prior experience transitioning traditional infrastructure teams into SRE practices
What Success Looks Like
  • Reduced incidents and faster recovery times
  • Measurable reduction in operational toil
  • Improved system availability and performance
  • Clear reliability metrics aligned with business outcomes
  • Engineering teams empowered to ship faster with confidence
Company Culture & Values

We are dedicated to fostering a collaborative, engaging, and inclusive environment and are committed to providing a workplace that empowers associates to be authentic and bring their best to work. We believe that associates do their best when they feel safe, understood, and valued, and we work diligently and collaboratively to ensure Broadridge is a company—and ultimately a community—that recognizes and celebrates everyone’s unique perspective.

Use of AI in Hiring As part of the recruiting process, Broadridge may use technology, including artificial intelligence (AI)-based tools, to help review and evaluate applications. These tools are used only to support our recruiters and hiring managers, and all employment decisions include human review to ensure fairness, accuracy, and compliance with applicable laws. Please note that honesty and transparency are critical to our hiring process. Any attempt to falsify, misrepresent, or disguise information in an application, resume, assessment, or interview will result in disqualification from consideration.

Broadridge Financial Solutions (NYSE: BR) is a global technology leader with trusted expertise and transformative technology, helping clients and the financial services industry operate, innovate, and grow. We power investing, governance, and communications for our clients - driving operational resiliency, elevating business performance, and transforming investor experiences. Our technology and operations platforms process and generate over 7 billion communications annually and underpin the daily average trading of over $15 trillion in equities, fixed income, and other securities globally. A certified Great Place to Work®, Broadridge is part of the S&P 500® Index, employing over 15,000 associates in 21 countries.

LinkedIn Facebook Instagram Twitter YouTube Glassdoor The Muse Broadridge is committed to creating an engaging workplace for the most talented associates in our industry. We are dedicated to fostering a collaborative, inclusive, and healthy environment that promotes flexibility and accountability. As a leading provider of technology, communications, and data and analytics solutions to businesses around the world, it is critical that we understand, embrace, and operate in a multicultural environment. Every associate has unique strengths, which, when fully appreciated and embraced, allow individuals to perform at their best, leading to our success. We believe that our associates are our most important asset. Encouraging professional development opportunities is a core part of our culture. Broadridge provides educational opportunities, including formal classes, training programs and events. To enable learning in our hybrid working model, Broadridge has redesigned all development programs for 100% virtual delivery. Our associates have access to 8,500+ online courses covering business, leadership, technical, and function-specific topics through our LinkedIn Learning program.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

Broadridge Financial Solutions • Manila, Hinoba-an

On-site
PHP 1,800,000 - 2,600,000
Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

broadridge • Philippines

Hybrid
PHP 1,000,000 - 2,400,000
Hybrid work model
Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

Broadridge Connectivity Solutions • Manila

Hybrid
PHP 1,500,000 - 2,300,000
Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

Broadridge • Metro Manila

On-site
PHP 1,200,000 - 2,400,000
Senior Technical Analyst (Hybrid)
Senior Technical Analyst (Hybrid)

Broadridge Financial Solutions • Philippines

On-site
PHP 600,000 - 1,000,000
Senior Technical Analyst (Hybrid)
Senior Technical Analyst (Hybrid)

Broadridge Financial Solutions • Manila, Hinoba-an

On-site
PHP 500,000 - 800,000
Site Reliability Engineer (Windows)
Site Reliability Engineer (Windows)

Broadridge • Metro Manila

On-site
PHP 1,116,000 - 2,232,000
Site Reliability Engineer (AWS)
Site Reliability Engineer (AWS)

Broadridge Connectivity Solutions • Manila

Hybrid
PHP 1,000,000 - 1,800,000
Process Analyst (Hybrid)
Process Analyst (Hybrid)

Broadridge Financial Solutions • Philippines

On-site
PHP 600,000 - 1,200,000
Process Analyst (Hybrid)
Process Analyst (Hybrid)

Broadridge Financial Solutions • Manila, Hinoba-an

On-site
PHP 500,000 - 800,000