Manager, Site Reliability Engineer

Forge

New York (NY)

On-site

USD 150,000 - 220,000

Full time

8 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Forge in New York seeks an experienced Manager, Site Reliability Engineering to lead our SRE team, drive reliability, observability, and incident response across the platform, and partner with Platform, Security, Compliance, and Product teams.

This hands-on technical leadership role emphasizes coaching engineers, improving production operations, and delivering secure, scalable services while aligning with Forge's mission to power a transparent, accessible private market.

Qualifications

  • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
  • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
  • Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
  • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
  • Hands-on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
  • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
  • Strong technical judgment, communication skills, and ability to influence across engineering and non-engineering stakeholders.

Responsibilities

  • Lead Forge's Site Reliability Engineering team to keep Forge systems highly available for customers.
  • Drive incident management practices with engineering teams, including response, mitigation, follow-up, and post-incident learning.
  • Build, improve, and manage observability infrastructure with Platform Engineering, including monitoring, alerting, dashboards, and metrics.
  • Improve monitoring coverage and alert quality to reduce noise and speed detection and response.
  • Champion reliability best practices across engineering, including service ownership and disaster recovery.
  • Contribute to technical design, architecture, automation, infrastructure, and team delivery.
  • Collaborate with engineering teams to troubleshoot production issues and improve system reliability.
  • Hire, coach, mentor, and manage performance for SRE team members.

Skills

SRE leadership
Cloud operations
Observability
CI/CD
Automation
Incident management

Education

Bachelor's degree in CS/Engineering

Tools

AWS
Azure
Kubernetes
Terraform
Datadog
CloudWatch

Job description

At Forge, we know our team is our greatest asset. As technology innovators in the private market, our vision is to deliver a richer future for everyone. We live that vision through our values of being bold, accountable, and humble. We experience the value that our vision brings to the world every day, helping the teams behind the greatest innovations of our generation, from space travel to artificial intelligence, and more.

With liquidity solutions, exclusive data and insights, a custody offering, and a vibrant marketplace, Forge's goal is to build the best-in-class technology infrastructure to power a global private market that is transparent, accessible, and seamless for companies, their employees, and investors. Through Forge, employees can sell their private shares, employers can reward shareholders with pre-IPO liquidity and individual and institutional investors can participate in private unicorn growth.

Forge's differentiated global marketplace addresses rising demand among individual and institutional investors for exposure to private company stocks and is building a growing network effect.

Our ability to offer these powerful financial solutions has generated incredible interest from investors, demand from customers, and a need to grow our team to meet the needs of more companies, teams, and innovators in this way.

The Role:

As an engineering organization, we pride ourselves on engineering as a creative activity. Engineering managers enable engineers to do their best work by maintaining a culture and environment where engineers can achieve autonomy, mastery, and purpose. The Manager, Site Reliability Engineering will lead Forge's SRE team responsible for keeping Forge systems highly available for customers, while partnering closely with Platform, Engineering, Security, Compliance, and Product teams to improve reliability, observability, incident response, and operational maturity. This is an opportunity for a hands‑on technical leader who can coach engineers, improve production operations, and help Forge build and run secure, scalable, and highly reliable products.

Responsibilities:
  • Manage Forge's Site Reliability Engineering team responsible for keeping Forge systems highly available for customers.
  • Drive strong incident management practices in partnership with engineering teams, including response, mitigation, follow‑up, and post‑incident learning.
  • Build, improve, and manage observability infrastructure in partnership with Platform Engineering, including monitoring, alerting, dashboards, and operational metrics.
  • Improve monitoring coverage and alert quality to reduce noise, shorten time to detect, and support faster response and mitigation.
  • Champion reliability best practices across engineering, including service ownership, operational readiness, disaster recovery, and production support standards.
  • Contribute to technical design, architecture, automation, infrastructure, and overall team delivery.
  • Collaborate with engineering teams to troubleshoot production issues, identify recurring problems, and improve system reliability.
    • Hire, coach, mentor, and manage performance for SRE team members while supporting career development and team health.
  • Partner with Security, Compliance, and Risk partners to ensure reliability and infrastructure practices meet the needs of a regulated business.
Qualifications:
  • 5+ years of experience leading a Site Reliability Engineering, DevOps, Cloud Operations, or similar reliability-focused function.
  • 10+ years of total software engineering, infrastructure, platform, cloud, or production operations experience.
  • Bachelor's degree in Computer Science, Engineering, or a closely related field, or equivalent practical experience.
  • Experience building, operating, and maintaining large-scale cloud infrastructure and distributed systems.
  • Hands‑on experience with observability, monitoring, alerting, incident response, troubleshooting, and production support.
  • Experience with CI/CD, infrastructure automation, cloud platforms, and operational tooling.
  • Strong technical judgment, communication skills, and ability to influence across engineering and non‑engineering stakeholders.
Preferred Qualifications:
  • Experience in FinTech, financial services, or another regulated industry.
  • Experience with AWS and/or Azure cloud platforms.
  • Familiarity with Kubernetes, container platforms, infrastructure‑as‑code, Terraform, Ansible, or similar automation tooling.
  • Experience with observability platforms such as Datadog, CloudWatch, or similar tools.
  • Experience improving developer experience through paved‑road platforms, standardization, and self‑service infrastructure capabilities.
  • Experience supporting growth‑stage companies where speed, scale, reliability, and operational discipline must be balanced.

For residents of San Francisco/Bay Area, CA or New York, NY the annual salary range for this role is $150,000-$220,000 + annual bonus. Final offers may vary from the amount listed based on geography, candidate experience and expertise, annual bonus, and other factors.

Upon offer, we conduct background checks that include employment and education verification, state, and county criminal history searches as well as fingerprint and drug test.

Forge is proud to be an equal opportunity employer committed to supporting a diverse and inclusive workplace. Our employment decisions are made without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), gender, gender identity, gender expression, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, marital status, sexual orientation, veteran status, or any other characteristic protected by federal, state, or local laws.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Manager, Site Reliability Engineer New San Francisco, California, United States
Manager, Site Reliability Engineer New San Francisco, California, United States

Forge Global • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 220,000
Manager, Site Reliability Engineer
Manager, Site Reliability Engineer

Socket.dev • New York (NY)

On-site
USD 150,000 - 220,000
Manager, Site Reliability Engineer
Manager, Site Reliability Engineer

Socket.dev • San Francisco (CA)

On-site
USD 150,000 - 220,000
Manager, Site Reliability Engineer New York, New York, United States
Manager, Site Reliability Engineer New York, New York, United States

Forge Global • New York (NY), Northern (KY)

Hybrid
USD 150,000 - 220,000
Director of Platform & Reliability Engineering
Director of Platform & Reliability Engineering

Forge Global • San Francisco (CA)

Hybrid
USD 235,000 - 245,000
Annual bonus
Diverse and inclusive workplace
Director of Software Engineering, Client Experience San Francisco, California, United States
Director of Software Engineering, Client Experience San Francisco, California, United States

Forge Global, Inc. • San Francisco (CA)

Hybrid
USD 245,000 - 270,000
Senior Manager, Data Engineering
Senior Manager, Data Engineering

Forge • San Francisco (CA)

Hybrid
USD 245,000 - 270,000
Senior Software Engineer (Tech Lead), Customer Domain Engineering
Senior Software Engineer (Tech Lead), Customer Domain Engineering

Forge • San Francisco (CA)

Hybrid
USD 209,000 - 240,000
Engineering Manager, Trade Management
Engineering Manager, Trade Management

Forge • New York (NY)

On-site
USD 225,000 - 250,000
Senior Software Engineer II of Trading
Senior Software Engineer II of Trading

Forge • San Francisco (CA)

Hybrid
USD 174,000 - 210,000