BXTI, Site Reliability Engineer - Data, Cloud & Developer Experience

InforCapital

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Blackstone is seeking a Site Reliability Engineer to lead the adoption of SRE practices, improve observability, and ensure reliable services across the firm. The role involves instrumentation, monitoring, and automation to reduce toil and incident impact.

The successful candidate will collaborate with development and operations teams to design resilient systems, participate in on-call rotations, and drive continuous improvement in reliability and performance.

Responsibilities

  • Provide technical leadership in the understanding and adoption of SRE methodologies across the firm.
  • Incorporating observability standards into code and deployment pipelines.
  • Evolving the SRE standards that are adopted across all teams.
  • Partnering with colleagues in various roles to improve service reliability and operational efficiency.
  • Assisting developers and engineers directly and through AI assistants.
  • Implement instrumentation and provide comprehensive performance insights to service owners.
  • Ensuring monitoring and alerting that reflects the reliability of services for users and enables effective on-call operations.
  • Implementing strategic observability tools and working to control overhead in maintenance and cost.
  • Participate in on-call rotations and respond to system incidents to ensure service availability and minimize operational impact.
  • Using automation to manage, maintain, and scale SRE systems with minimal human intervention.
  • Fostering a culture of reliability and continuous improvement across product engineering teams.

Job description

LocationLondon, United Kingdom## About This RoleBlackstone is the world’s largest alternative asset manager. We seek to create positive economic impact and long-term value for our investors, the companies we invest in, and the communities in which we work. We do this by using extraordinary people and flexible capital to help companies solve problems. Our $1.1 trillion in assets under management include investment vehicles focused on private equity, real estate, public debt and equity, infrastructure, life sciences, growth equity, opportunistic, non-investment grade credit, real assets and secondary funds, all on a global basis. Further information is available at www.blackstone.com. Follow @blackstone on LinkedIn, X, and Instagram. Blackstone’s Site Reliability Engineering team is responsible for improving the reliability of systems and services to meet the needs of the business. This is achieved through collaboration with the development and engineering teams to leverage SRE practices and principles. You’ll have the opportunity to identify and solve new problems as they arise, deploy and maintain observability systems and pipelines, mature the operations and support of services and platforms, and pursue emerging opportunities for efficiency and business value. This position involves the selection, implementation, and maintenance of key observability tooling. It requires ongoing evaluation of the firm’s needs in observability, monitoring, alerting, resilience, and recovery. We work alongside service owners on design, implementation, and management of services for continuous improvement. We achieve the requisite reliability of services using clear definitions and measurable targets. We plan for and practice recovery from disaster scenarios and respond in real time to incidents. We guide the postmortem process in order to mitigate risks, prevent future disruptions, and improve the on-call experience. We aim to eliminate manual work, improve operational efficiency, and ensure the high quality outputs in all that we do. Key Responsibilities: Provide technical leadership in the understanding and adoption of SRE methodologies across the firm Incorporating observability standards into code and deployment pipelines. Evolving the SRE standards that are adopted across all teams Partnering with colleagues in various roles and reporting lines to improve service reliability and operational efficiency Assisting developers and engineers directly and through AI assistants. Implement instrumentation and provide comprehensive performance insights to service owners Ensuring monitoring and alerting that reflects the reliability of services for users and enables effective on-call operations Implementing strategic observability tools and working to control overhead in maintenance and cost Participate in on-call rotations and respond to system incidents to ensure service availability and minimize operational impact Using automation to manage, maintain, and scale SRE systems with minimal human intervention Fostering ...
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Lead: Data, Cloud & Developer Experience
SRE Lead: Data, Cloud & Developer Experience

InforCapital • Greater London

Hybrid
GBP 90,000 - 130,000
SRE
SRE

Technopride Ltd • Hove

On-site
GBP 60,000 - 80,000
Site Reliability Engineer (SRE) – Cloud Platforms
Site Reliability Engineer (SRE) – Cloud Platforms

Talenzon group • Greater London

On-site
GBP 70,000 - 110,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000
Observability SRE
Observability SRE

HCLTech • Greater London

On-site
GBP 70,000 - 95,000
Principal Site Reliability Engineer, Infrastructure Observability
Principal Site Reliability Engineer, Infrastructure Observability

United States Digital Space LLC • Greater London

On-site
GBP 120,000 - 170,000
Hybrid work up to 3 days per week
Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Sutton

On-site
GBP 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

SR2 | Socially Responsible Recruitment | Certified B Corporation • Slough

On-site
GBP 65,000 - 90,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000
Lead SRE - Charing Cross
Lead SRE - Charing Cross

Hackajob Ltd • Leeds

On-site
GBP 61,000 - 101,000