Global Manager of Site Reliability Engineering (Hybrid - Flexible Options)

Broadridge Financial Solutions

New York (NY)

Hybrid

USD 235,000 - 250,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Broadridge is seeking a strategic, hands‑on Global Manager of Site Reliability Engineering to lead reliability, release engineering, and operational evolution of its Integrated Platform. You will build and develop a high‑performing SRE team while staying deeply engaged in architecture, software engineering, automation, and production performance.

You will partner with application engineering, product, architecture, security, and operations to ensure reliability is engineered into services from

Qualifications

  • 10+ years of software or production engineering experience with Java-based services.
  • Experience leading engineers and managing delivery and accountability.
  • Hands-on AWS experience with IAM, networking, observability and infrastructure as code.
  • Strong Java and Spring Boot skills, API design, concurrency, performance, and secure coding.

Responsibilities

  • Build and lead a high-performing SRE team.
  • Own SRE strategy and execution roadmap.
  • Improve CI/CD pipelines through automation and testing.
  • Ensure operational reliability, performance, and production readiness.
  • Lead incident resolution and resilience improvements.

Skills

Java
Spring Boot
AWS
Kafka
PostgreSQL
People management
Automation
CI/CD
Architecture
Incident response

Tools

Kubernetes
Terraform
OpenTelemetry
Kafka Connect
PostgreSQL HA

Job description

At Broadridge, we've built a culture where the highest goal is to empower others to accomplish more. If you’re passionate about developing your career, while helping others along the way, come join the Broadridge team. Broadridge is Growing. We are seeking a strategic, hands‑on Global Manager of Site Reliability Engineering to lead the reliability, release engineering, and operational evolution of its Integrated Platform. This event‑driven platform connects enterprise applications with client‑facing experiences through a common ontology, standardized APIs, and scalable event‑driven capabilities across Capital Markets, Wealth Management, Investor Communications, and other business domains. This leader will build and develop a high‑performing SRE team while remaining deeply engaged in architecture, software engineering, automation, and production performance. Partnering with application engineering, product, architecture, security, and operations, the Senior Manager will ensure that reliability is engineered into services from design through deployment and ongoing operation. The role requires strong technical depth in Java, Spring Boot, AWS, Kafka, and PostgreSQL, combined with demonstrated people leadership and delivery accountability. DevOps practices, release automation, and responsible AI enablement will be central to improving engineering productivity and production outcomes.

Job responsibilities:
  • Build and lead a high-performing SRE team.
  • Own hiring, coaching, performance management, career development, and succession planning.
  • Establish clear technical expectations, strengthen engineering judgment, and maintain sustainable operational and on‑call responsibilities.
  • Own the SRE strategy and execution roadmap.
  • Translate platform and business priorities into measurable improvements in reliability, release performance, automation, and scalability.
  • Manage capacity, dependencies, and delivery commitments while protecting time for engineering improvements.
  • Remain hands‑on in engineering decisions.
  • Lead architecture and code reviews, guide complex troubleshooting, and contribute to critical automation and tooling.
  • Ensure operational software meets the same standards for testing, security, maintainability, and documentation as application code.
  • Advance AWS infrastructure and platform automation.
  • Develop reusable infrastructure‑as‑code patterns, consistent environment configurations, and automated provisioning and recovery.
  • Strengthen IAM, networking, compute resilience, and capacity management across the platform’s AWS services.
  • Improve Kafka reliability and event‑processing performance.
  • Address partitioning, consumer‑group behavior, consumer lag, schema evolution, delivery semantics, replay, and failure recovery.
  • Validate that recovery approaches preserve intended processing behavior and data integrity.
  • Strengthen PostgreSQL performance and resilience.
  • Partner with application and database engineers on schema design, indexing, query optimisation, transaction behavior, connection pooling, safe migrations, and recovery testing.
  • Lead release engineering and deployment readiness.
  • Improve CI/CD pipelines through automated testing, security checks, artifact traceability, and production validation.
  • Establish safe deployment and rollback or roll‑forward patterns that account for API compatibility, database changes, and event‑schema dependencies.
  • Make service health measurable.
  • Define service‑level indicators, service‑level objectives, and error‑budget practices with service owners.
  • Use metrics, logs, and distributed traces to improve detection, reduce alert noise, and prioritise engineering work based on business impact.
  • Reduce operational toil through automation.
  • Build reusable tooling, self‑service capabilities, and controlled remediation for well‑understood failure scenarios.
  • Measure reductions in manual effort, recurring incidents, and recovery time.
  • Enable practical, responsible AI adoption.
  • Integrate approved AI tools into code and test development, infrastructure review, knowledge retrieval, and incident investigation.
  • Validate outputs, measure benefits, protect sensitive information, and retain appropriate human review and authorisation for production changes.
  • Lead incident resolution and resilience improvement.
  • Coordinate technical response to complex incidents, communicate impact and recovery progress, and drive blameless reviews and permanent corrective actions.
  • Validate capacity, failover, backup restoration, and disaster recovery against agreed objectives.
  • Influence cross‑functional decisions.
  • Communicate technical risk and investment trade‑offs to senior stakeholders.
  • Resolve competing priorities and establish shared accountability for production readiness, secure delivery, and reliable service operation.
Leadership and professional capabilities

Functional knowledge: Deep software and reliability engineering expertise, with sufficient breadth across cloud infrastructure, data platforms, security, and operations to guide end‑to‑end technical decisions.

Business expertise: Understand how platform failures and delivery constraints affect clients and business services. Anticipate operational and regulatory risks and translate them into practical engineering priorities.

Leadership: Deliver through others without becoming detached from the technology. Develop technical leaders, delegate meaningful ownership, and hold teams accountable for engineering quality and measurable outcomes.

Problem solving: Resolve complex, cross‑service problems using evidence, experimentation, and sound technical judgement. Favor durable engineering solutions over repeated manual intervention.

Interpersonal skills: Explain complex technical issues clearly, constructively challenge assumptions, and influence senior stakeholders across organizational boundaries.

Required qualifications:
  • 10+ years of software engineering or closely related production engineering experience, including substantial experience building and operating Java‑based services in production.
  • Demonstrated engineering people‑management experience, including hiring, performance management, talent development, resource planning, and accountability for complex technical delivery.
  • Experience leading senior engineers and developing technical leads.
  • Strong Java and Spring Boot skills, including API design, concurrency, performance tuning, automated testing, and secure coding.
  • Ability to review application code and diagnose runtime behaviour rather than relying solely on infrastructure‑level investigation.
  • Hands‑on AWS experience deploying and operating services, with knowledge of IAM, networking, observability, infrastructure as code, and at least one compute platform such as ECS, EKS, or Lambda.
  • Production Kafka experience covering event‑driven design, consumer groups, partitioning, schema evolution, delivery semantics, and failure recovery.
  • Strong PostgreSQL skills in schema design, indexing, query optimisation, transactions, migrations, and operational troubleshooting.
  • Demonstrated ability to build maintainable automation and operational tooling using sound software engineering practices, including version control, peer review, automated testing, and reusable design.
  • Experience with CI/CD, release engineering, automated testing, monitoring, incident response, and reliable distributed‑system design.
  • Practical understanding of service‑level objectives, capacity planning, production readiness, recovery strategies, and the use of delivery and reliability measures to guide improvement.
  • Ability to lead technical design reviews, mentor engineers, and communicate architectural, operational, and delivery trade‑offs with product, security, operations, and senior technology leaders.
Preferred qualifications
  • Experience with Kubernetes, Terraform, OpenTelemetry, Kafka Connect, and PostgreSQL replication or high availability.
  • Experience designing and operating platforms with clearly defined service‑level objectives, capacity plans, and tested disaster‑recovery procedures.
  • Experience implementing progressive delivery, automated release verification, self‑service engineering tools, or policy‑as‑code controls.
  • Experience introducing AI‑assisted engineering or operational capabilities with demonstrated improvements in productivity, quality, or incident resolution, supported by appropriate evaluation and safeguards.
  • Experience supporting financial‑services platforms or other regulated, business‑critical distributed systems.
Salary Range

235,000.00 – 250,000.00 USD annual Bonus Eligible Please visit www.broadridgebenefits.com for more information on our comprehensive benefit offerings. #LI-MR1 #LI-Hybrid

Broadridge considers various factors when evaluating a candidate's final salary including, but not limited to, relevant experience, skills, and education.

Use of AI in Hiring

As part of the recruiting process, Broadridge may use technology, including artificial intelligence (AI)-based tools, to help review and evaluate applications. These tools are used only to support our recruiters and hiring managers, and all employment decisions include human review to ensure fairness, accuracy, and compliance with applicable laws. Please note that honesty and transparency are critical to our hiring process. Any attempt to falsify, misrepresent, or disguise information in an application, resume, assessment, or interview will result in disqualification from consideration.

Disability Assistance

We recognize that ensuring our long‑term success means creating an environment where everyone is welcome, where everyone's strengths are valued, and where everyone can perform at their best.

Broadridge provides equal employment opportunities to all associates and applicants for employment without regard to race, color, religion, sex (including sexual orientation, gender identity or expression, and pregnancy), marital status, national origin, ethnic origin, age, disability, genetic information, military or veteran status, and other protected characteristics protected by applicable federal, state, or local laws.

If you need assistance or would like to request reasonable accommodations during the application and/or hiring process, please contact us at 888‑237‑7769 or by sending an email to BRcareers@broadridge.com.

Broadridge Financial Solutions (NYSE: BR) is a global technology leader with trusted expertise and transformative technology, helping clients and the financial services industry operate, innovate, and grow. We power investing, governance, and communications for our clients – driving operational resiliency, elevating business performance, and transforming investor experiences. Our technology and operations platforms process and generate over 7 billion communications annually and underpin the daily average trading of over $15 trillion in equities, fixed income, and other securities globally. A certified Great Place to Work®, Broadridge is part of the S&P 500® Index, employing over 15,000 associates in 21 countries. LinkedIn Facebook Instagram Twitter YouTube Glassdoor The Muse

Broadridge is committed to creating an engaging workplace for the most talented associates in our industry. We are dedicated to fostering a collaborative, inclusive, and healthy environment that promotes flexibility and accountability.

As a leading provider of technology, communications, and data and analytics solutions to businesses around the world, it is critical that we understand, embrace, and operate in a multicultural environment. Every associate has unique strengths, which, when fully appreciated and embraced, allow individuals to perform at their best, leading to our success. We believe that our associates are our most important asset. Encouraging professional development opportunities is a core part of our culture.

Broadridge provides educational opportunities, including formal classes, training programs and events. To enable learning in our hybrid working model, Broadridge has redesigned all development programs for 100% virtual delivery. Our associates have access to 8,500+ online courses covering business, leadership, technical, and function‑specific topics through our LinkedIn Learning program.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Software Engineer (Hybrid)
Lead Software Engineer (Hybrid)

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 150,000 - 170,000
Senior Software Engineer (Hybrid - Newark, NJ or NYC)
Senior Software Engineer (Hybrid - Newark, NJ or NYC)

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 120,000 - 130,000
VP Software Engineering
VP Software Engineering

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 255,000 - 265,000
Senior Software Engineer (Hybrid - Edgewood, NY)
Senior Software Engineer (Hybrid - Edgewood, NY)

Broadridge Financial Solutions • Edgewood (NY)

Hybrid
USD 120,000 - 130,000
VP, Head of Solutions Engineers (Hybrid / NYC)
VP, Head of Solutions Engineers (Hybrid / NYC)

Broadridge Financial Solutions • Town of Florida (NY)

Hybrid
USD 200,000 - 220,000
Bonus Eligible
Benefits information
Senior Process Improvement Analyst (Hybrid)
Senior Process Improvement Analyst (Hybrid)

Broadridge Financial Solutions • New York (NY)

On-site
USD 115,000 - 120,000
Bonus Eligible
Benefits Information
Sr Software Engineer (Hybrid - Newark, NJ)
Sr Software Engineer (Hybrid - Newark, NJ)

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 120,000 - 130,000
Bonus Eligible
Employee Benefits
Software Development Manager (SDET & Test Automation) (Hybrid - Newark, NJ or NYC)
Software Development Manager (SDET & Test Automation) (Hybrid - Newark, NJ or NYC)

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 200,000 - 215,000
Benefits information
Paid sick leave (Colorado)
Lead Financial Analyst Corporate Technology- Infrastructure (Hybrid)
Lead Financial Analyst Corporate Technology- Infrastructure (Hybrid)

Broadridge Financial Solutions • Newark (NJ)

On-site
USD 125,000 - 130,000
Bonus Eligible
VP, Product, Digital Asset Platform
VP, Product, Digital Asset Platform

Broadridge Financial Solutions • New York (NY)

On-site
USD 200,000 - 260,000