Platform Site Reliability Engineer (SRE)

Broadridge Connectivity Solutions

Manila

On-site

PHP 1,000,000 - 1,600,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Broadridge Connectivity Solutions in Manila is seeking a Platform Site Reliability Engineer to ensure stability, scalability, and reliability for our products. You will collaborate with EP teams to deliver robust platforms and drive NFRs across new products.

You will lead incident responses, build automation, and guide the team in continuous improvement while coordinating with cross-functional groups within the Enterprise Platform.

Qualifications

  • Bachelor's degree in CS/IT/SE or related field.
  • 5+ years in production applications support (SRE/DevOps).
  • Experience with SLOs and SLIs and incident response.
  • Proficient in Windows and/or Linux administration.
  • Experience with observability tools (Datadog, Splunk).
  • Ability to automate ops with Python/Java/Shell/Terraform/Ansible.
  • Experience with AWS and middleware like databases, MQ, Kafka.
  • Familiar with Docker and Kubernetes; strong problem solving.

Responsibilities

  • Implement SRE best practices, SLOs/SLIs, and monitoring.
  • Build automation to improve reliability of platforms.
  • Perform capacity planning and design for load growth.
  • Troubleshoot incidents and perform root cause analysis.
  • Participate in incident calls for outages and emergencies.
  • Collaborate with developers to define reliability requirements.
  • Conduct post-mortems and drive improvements.
  • Use data to prevent incidents and optimize stability.
  • Support product teams to improve stability and extensibility.
  • Drive standard implementation of NFRs in new products.

Skills

SRE/DevOps
SLOs/SLIs
Windows/Linux Admin
Observability
Automation & Scripting
AWS
Middleware & Kafka
Docker/Kubernetes
English Communication
Team Collaboration
Continual Improvement

Education

BSc in CS/IT/SE

Tools

Datadog
Splunk
Terraform
Chef
Puppet
Ansible
SQL
Docker
Kubernetes

Job description

At Broadridge, we've built a culture where the highest goal is to empower others to accomplish more. If you’re passionate about developing your career, while helping others along the way, come join the Broadridge team.

Role Overview

As a Platform Site Reliability Engineer (SRE), you will play a critical role in ensuring the stability, scalability, and reliability of our products and services. You will work closely with cross-functional teams to design, develop, and deploy solutions that enhance the performance and uptime of our applications.

The Platform SRE is part of the Enterprise Platform (EP) group and is responsible for supporting and running our standard platforms efficiently and effectively. You will be expected to collaborate closely with other functions within EP (DevOps/Cloud Platforms, Quality Engineering and Developer Experience) to provide robust, integrated and best-in-class solutions for our product engineering teams.

Key Responsibilities
  • Implementing Site Reliability Engineering best practices, including error budgeting, service level objectives (SLOs), and monitoring and alerting systems

  • Building automation tools and processes to improve the efficiency and reliability of running our products and standard platforms

  • Performing capacity planning and system design to ensure that our systems can handle increasing traffic and load

  • Troubleshooting complex technical issues and providing root cause analysis to prevent future incidents

  • Participating in incident calls to respond to system outages and emergencies

  • Collaborating with software developers to define and implement reliability requirements for new products/ applications/services

  • Conducting post-mortem analyses to identity opportunities for improvement and prevent recurring issues

  • Using data-based decision making to be proactive in the prevention of potential incidents and problems

  • Supporting product development teams in the implementation of tools, processes, and practices to improve stability, reliability, and extensibility of their products

  • Collaborate across the EP function to ensure that standard platforms are best-in-class

  • Drive standard implementation of NFRs in new product development and own the "deep-dive" process to improve problematic application

  • Overall management and governance of vulnerabilities and End-of-life within our products

As a senior member of the team, you will be responsible for:

  • Overseeing initiatives and deliverables across the team

  • Technical and design decisions made by the team.

  • Coaching and mentoring members of the team

  • Keeping your finger on the pulse: identifying and developing new ideas and initiatives

  • Acting as an advocate for SRE across Enterprise Platform team and wider Broadridge community

  • Reviewing work and improving SRE processes

  • Collaborating with other teams outside Enterprise
    Platforms

  • Contributing to the strategic direction of the function

Skills And Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Software Engineering, or a related field.

  • 5+ years of experience supporting production applications (ie SRE and/or DevOps roles)

  • Practical understanding of implementing SLOs and SLIs

  • Knowledge of Windows and/or Linux Systems administration and networking fundamentals

  • Experience in implementing Observability and Alerting tools (eg Datadog, Splunk)

  • Ability to automate application operations using tools such as Python, Java, Shell Scripting, Terraform, Chef, Puppet, SQL, Ansible

  • Knowledge of AWS

  • Experience in supporting middleware such as databases, webservers, MQ and Kafka

  • Familiarity with containerization technologies, such as Docker and Kubernetes

  • Excellent problem-solving skills and attention to detail

  • Ability to work well under pressure and prioritize tasks in a fast-paced environment

  • Fluency in English is essential.

  • Ability to collaborate closely with others.

  • Continual Improvement mindset

We are dedicated to fostering a collaborative, engaging, and inclusive environment and are committed to providing a workplace that empowers associates to be authentic and bring their best to work. We believe that associates do their best when they feel safe, understood, and valued, and we work diligently and collaboratively to ensure Broadridge is a company—and ultimately a community—that recognizes and celebrates everyone’s unique perspective.

Use of AI in Hiring

As part of the recruiting process, Broadridge may use technology, including artificial intelligence (AI)-based tools, to help review and evaluate applications. These tools are used only to support our recruiters and hiring managers, and all employment decisions include human review to ensure fairness, accuracy, and compliance with applicable laws. Please note that honesty and transparency are critical to our hiring process. Any attempt to falsify, misrepresent, or disguise information in an application, resume, assessment, or interview will result in disqualification from consideration.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Platform Site Reliability Engineer (SRE)
Platform Site Reliability Engineer (SRE)

Broadridge Financial Solutions • Manila, Hinoba-an

Hybrid
PHP 1,200,000 - 1,800,000
Platform Site Reliability Engineer (SRE)
Platform Site Reliability Engineer (SRE)

Broadridge Financial Solutions • Philippines

On-site
PHP 1,000,000 - 1,200,000
Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

Broadridge • Metro Manila

On-site
PHP 1,200,000 - 2,400,000
Site Reliability Engineer (Windows)
Site Reliability Engineer (Windows)

Broadridge • Metro Manila

On-site
PHP 1,116,000 - 2,232,000
Senior Site Reliability Engineer (AWS)
Senior Site Reliability Engineer (AWS)

Broadridge Financial Solutions • Manila, Hinoba-an

On-site
PHP 1,800,000 - 2,600,000
Site Reliability Engineer (AWS)
Site Reliability Engineer (AWS)

Broadridge • Metro Manila

Hybrid
PHP 1,000,000 - 1,800,000
Hybrid work arrangement
Onsite 4x per month in Makati City
Senior Site Reliability Engineer (Windows)
Senior Site Reliability Engineer (Windows)

Broadridge • Metro Manila

On-site
PHP 1,200,000 - 2,100,000
Platform SRE: Reliability & Automation Architect
Platform SRE: Reliability & Automation Architect

Broadridge Financial Solutions • Philippines

Hybrid
PHP 1,000,000 - 1,200,000
Platform SRE: Scale, Automate & Reliability (Hybrid)
Platform SRE: Scale, Automate & Reliability (Hybrid)

Broadridge Financial Solutions • Manila, Hinoba-an

Hybrid
PHP 1,200,000 - 1,800,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG • Philippines

On-site
PHP 1,000,000 - 1,600,000