Staff Software Engineer, Site Reliability Engineer

Harvey, Inc.

San Francisco, Northern (CA, KY)

Hybrid

USD 238,000 - 290,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Relocation assistance
In-person work model

Job summary

Harvey is seeking a Staff Software Engineer for the Site Reliability team in San Francisco to ensure reliability, scalability, and performance of the global platform. You will own systems that keep the platform fast, secure, and always on.

You’ll design monitoring, lead incident response, automate operations, and mentor engineers. The role requires 10+ years in SRE, strong IaC experience with Pulumi/Terraform/CloudFormation, and deep cloud knowledge.

Qualifications

  • 10+ years in Site Reliability Engineering or similar roles.
  • Expertise in IaC tools (Pulumi, Terraform, CloudFormation).
  • Deep familiarity with observability tools (Datadog, Sentry) and incident response (PagerDuty).
  • Proficiency with cloud platforms (Azure, GCP, AWS).
  • Strong programming skills (Python, Bash, Go).
  • Understanding of CI/CD, Kubernetes, containerization, and cloud security.

Responsibilities

  • Design, implement, and manage monitoring, alerting, and infrastructure resources across 50+ regions.
  • Lead incident management processes, postmortems, and RCA analyses.
  • Automate operational tasks and workflows for capacity planning and safe data access.
  • Establish security, compliance, and reliability best practices across the lifecycle.
  • Optimize infrastructure costs through capacity planning and build-versus-buy decisions.
  • Provide technical mentorship and leadership across teams.

Skills

Mentorship
Programming (Python/Bash/Go)
CI/CD
Kubernetes
Observability
Cloud platforms (Azure/GCP/AWS)
Infrastructure as Code

Tools

Pulumi
Terraform
CloudFormation
Datadog
Sentry
PagerDuty

Job description

Why Harvey

At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come.


This is a rare chance to help build a generational company at a true inflection point. We have strong product-market fit and world-class investor support. We’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched.


Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you.


At Harvey, the future of professional services is being written today — and we’re just getting started.


Role Overview

As a Staff Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow. If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you.


This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees.


What You’ll Do


  • Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions


  • Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements


  • Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention


  • Establish best practices for security, compliance, and reliability and collaborate across teams to drive these principles throughout the software lifecycle


  • Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality


  • Provide technical mentorship and leadership, promoting best practices and fostering team growth



What You Have


  • 10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams


  • Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.)


  • Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.)


  • Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.)


  • Strong programming skills (Python, Bash, Go, or similar languages)


  • Proven track record of diagnosing complex system problems and implementing durable solutions


  • Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles


  • Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence



Compensation Range

$238,000 - $290,000 USD


Depending on your location, an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices [here].

Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law.


We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing accommodations@harvey.ai

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Site Reliability Engineer
Senior Software Engineer, Site Reliability Engineer

Harvey, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 200,000 - 260,000
Staff Software Engineer, Site Reliability Engineer (SRE)
Staff Software Engineer, Site Reliability Engineer (SRE)

Harvey • San Francisco (CA)

On-site
USD 250,000 - 290,000
Comprehensive health coverage
401k match up to 4%
Flexible PTO
+1
Staff Software Engineer, Full Stack
Staff Software Engineer, Full Stack

Showcify, Inc. • United States

Remote
USD 231,000 - 314,000
Senior Software Engineer, Site Reliability Engineer
Senior Software Engineer, Site Reliability Engineer

Harvey • San Francisco (CA)

On-site
USD 200,000 - 260,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Engg • New York (NY)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Engg • San Francisco (CA)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Harvey • New York (NY)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Harvey • San Francisco (CA)

Hybrid
USD 231,000 - 340,000
Staff Software Engineer, Full Stack
Staff Software Engineer, Full Stack

Harvey, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 231,000 - 314,000
Relocation assistance
Senior Software Engineer, Frontend
Senior Software Engineer, Frontend

Harvey • New York (NY)

On-site
USD 193,000 - 290,000