Staff Software Engineer, Site Reliability Engineer (SRE)

Harvey

San Francisco (CA)

On-site

USD 250,000 - 290,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Comprehensive health coverage
401k match up to 4%
Flexible PTO
Relocation assistance

Job summary

Harvey is seeking a Staff Software Engineer for the Site Reliability team in San Francisco. You will ensure reliability, scalability, and performance of our legal AI platform, owning systems that keep the service fast, secure, and always on.

This role sits at the intersection of infrastructure and product, with a focus on automation and resilience. You will lead across 50+ regions, mentor engineers, and drive best practices in security, observability, and reliability while collaborating with

Qualifications

  • 10+ years in Site Reliability Engineering or similar roles.
  • Expertise with IaC tools and cloud platforms.
  • Strong observability and incident response experience.
  • Proven ability to mentor technical teams.

Responsibilities

  • Design and manage monitoring, alerting, and infra across regions.
  • Lead incident management, postmortems, RCA, improvements.
  • Automate tasks for capacity planning and safe data access.
  • Establish security and reliability best practices across lifecycle.
  • Mentor teams and drive operational excellence.

Skills

SRE experience
IaC tools
Observability
Cloud platforms
Programming
Root cause analysis
CI/CD / Kubernetes
Networking & security

Tools

Pulumi
Terraform
CloudFormation
Datadog
Sentry
PagerDuty
IncidentIO
Kubernetes
Docker

Job description

Why Harvey

Harvey is a secure AI platform for legal and professional services that augments productivity and automates complex workflows. Harvey uses algorithms with reasoning-adept LLMs that have been customized and developed by our expert team of lawyers, engineers and research scientists. We’ve found product market fit and are scaling our team very quickly. Some reasons to join Harvey are:

  • Exceptional product market fit: We have partnered with the largest law firms and professional service providers in the world, including Paul Weiss, A&O Shearman, Ashurst, O’Melveny & Myers, PwC, KKR, and many others.

  • Strategic investors: Raised over $500 million from strategic investors including Sequoia, Google Ventures, Kleiner Perkins, and OpenAI.

  • World-class team: Harvey is hiring the best talent from DeepMind, Google Brain, Stripe, FAIR, Tesla Autopilot, Glean, Superhuman, Figma, and more.

  • Partnerships: Our engineers and researchers work directly with OpenAI to build the future of generative AI and redefine professional services.

  • Performance: 4x ARR in 2024.

  • Competitive compensation.

Role Overview

As a Staff Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow. If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you.

This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees.

What You’ll Do
  • Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions

  • Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements

  • Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention

  • Establish best practices for security, compliance, and reliability and collaborate across teams to drive these principles throughout the software lifecycle

  • Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality

  • Provide technical mentorship and leadership, promoting best practices and fostering team growth

What You Have
  • 10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams

  • Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.)

  • Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.)

  • Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.)

  • Strong programming skills (Python, Bash, Go, or similar languages)

  • Proven track record of diagnosing complex system problems and implementing durable solutions

  • Solid understanding of CI/CD, Kubernetes, containerization, networking, and cloud security principles

  • Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence

Compensation Range

$250,000 - $290,000 USD

Please find our CA applicant privacy notice here.

Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law.

We are in the early innings of a generational company. Joining early at a hypergrowth startup has proven to lead to exponential growth in responsibility, access, and ability.

Compensation

$250K - $290K Offers Equity

Additionally, this role is eligible to participate in our equity plan and benefits program. Benefits include, but not limited to: Comprehensive health, dental and vision coverage, retirement benefits (401k match up to 4%), and flexible PTO.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Site Reliability Engineer
Senior Software Engineer, Site Reliability Engineer

Harvey • San Francisco (CA)

On-site
USD 200,000 - 260,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Engg • San Francisco (CA)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Engg • New York (NY)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Harvey • New York (NY)

Hybrid
USD 231,000 - 340,000
Staff/Sr. Staff Software Engineer, Product Engineering
Staff/Sr. Staff Software Engineer, Product Engineering

Harvey • San Francisco (CA)

Hybrid
USD 231,000 - 340,000
Staff Corporate Security Engineer
Staff Corporate Security Engineer

Harvey • San Francisco (CA)

On-site
USD 220,000 - 330,000
Senior Software Engineer, Frontend
Senior Software Engineer, Frontend

Harvey • New York (NY)

On-site
USD 193,000 - 290,000
Staff Software Engineer, Frontend
Staff Software Engineer, Frontend

Harvey • New York (NY)

On-site
USD 231,000 - 340,000
Senior Software Engineer, Security Harvey AI San Francisco $220,000 - $330,000/yr
Senior Software Engineer, Security Harvey AI San Francisco $220,000 - $330,000/yr

Neura Market • San Francisco (CA)

On-site
USD 220,000 - 330,000
Staff Product Manager, Integrations
Staff Product Manager, Integrations

Harvey • New York (NY)

On-site
USD 220,000 - 260,000