Staff Software Engineer, Core Infrastructure

Harvey

San Francisco (CA)

On-site

USD 236,000 - 290,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Harvey is seeking a Staff Software Engineer for the Core Infrastructure team in San Francisco. The role involves designing and building scalable infrastructure systems that power our legal AI platform, ensuring both operational excellence and innovation. The ideal candidate has over 10 years of experience in Infrastructure Engineering, proficiency in cloud platforms such as Azure, and strong skills in Infrastructure as Code tools like Terraform. The compensation range is $236,000 - $290,000 USD.

Qualifications

  • 10+ years of experience in Infrastructure Engineering or Platform Engineering in a production environment.
  • Long track record building and scaling complex, large-scale distributed systems.
  • Deep proficiency with cloud infrastructure platforms (Azure preferred).
  • Strong fluency in Infrastructure as Code (IaC) tools.
  • Solid understanding of Kubernetes and cloud security.

Responsibilities

  • Design and build scalable, fault-tolerant infrastructure systems.
  • Lead technical initiatives for observability and incident response.
  • Architect and optimize distributed systems for reliability.
  • Mentor engineers and raise the technical bar across the organization.
  • Drive infrastructure-as-code practices using tools like Terraform.

Skills

Infrastructure Engineering
Platform Engineering
Distributed Systems
Cloud Infrastructure
Infrastructure as Code
Kubernetes
Python

Tools

Terraform
Pulumi
Datadog

Job description

Role Overview

As a Staff Software Engineer on the Core Infrastructure team at Harvey, you will design and build new infrastructure systems while simultaneously scaling and strengthening our existing infrastructure. Our infrastructure powers every user interaction with Harvey – processing billions of prompt tokens and millions of daily requests across our global legal AI platform.

You will work in an environment balanced between innovation – building new systems – and operational excellence, ensuring that Harvey remains resilient and efficient as it scales products, regions, customers, and usage. Your contributions will directly impact the reliability, scalability, and security of our platform as we serve the world's leading law firms and professional service providers.

This role is based in San Francisco, CA. We use an in‑person work model and offer relocation assistance to new employees.

What You’ll Do
  • Design and build scalable, fault‑tolerant infrastructure systems that power Harvey's AI platform across multiple cloud regions
  • Own and evolve our multi‑cloud infrastructure (Azure, GCP), including Kubernetes orchestration, networking, and container management
  • Lead technical initiatives around observability, incident response, and operational excellence – building systems that enable rapid detection and resolution of issues
  • Architect and optimize our distributed systems for reliability, including load balancing, quota management, and failover mechanisms
  • Partner with Product Engineering and Security teams to ensure our infrastructure is an accelerant, not a constraint
  • Drive infrastructure‑as‑code practices using tools like Terraform and Pulumi to enable reproducible, auditable deployments
  • Mentor engineers and raise the technical bar across the organization through code reviews, design reviews, and technical leadership
Representative Projects
  • Design and implement a next‑generation model proxy architecture that routes millions of daily inference requests while maintaining model API compatibility and enabling seamless model integration
  • Build distributed rate‑limiting and quota management systems using Redis‑backed algorithms to handle bursty traffic patterns without degrading user experience
  • Architect multi‑region deployment strategies that meet strict data residency requirements for global enterprise customers
  • Develop comprehensive observability infrastructure with granular SLA monitoring, burn‑rate alerts, and detailed token attribution for cost tracking
  • Lead the evolution of our CI/CD pipelines to improve developer velocity while maintaining production stability
What You Have
  • 10+ years of experience in Infrastructure Engineering or Platform Engineering in a production environment
  • Long track record building and scaling complex, large‑scale distributed systems
  • Deep proficiency with cloud infrastructure platforms (Azure preferred; GCP or AWS experience transfers well)
  • Strong fluency in Infrastructure as Code (IaC) tools – Terraform, Pulumi, or CloudFormation
  • Solid understanding of Kubernetes, container orchestration, networking, and cloud security at scale
  • Experience with observability tools (Datadog, Sentry) and incident response practices (PagerDuty, Incident.io)
  • Strong programming skills in Python, Go, or similar languages
  • Excellent problem‑solving skills, a "spidey sense" of where things could go wrong, and a commitment to operational excellence
Nice to Have
  • Experience building infrastructure for AI/ML workloads or high‑throughput inference systems
  • Background with distributed rate‑limiting, load balancing, or quota management systems
  • Experience operating multi‑tenant platforms with strict security and compliance requirements
  • Track record of leading complex cross‑functional projects and delivering measurable impact
Compensation Range

$236,000 - $290,000 USD

Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing accommodations@harvey.ai

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer, Core Infrastructure
Staff Software Engineer, Core Infrastructure

Harvey • New York (NY)

On-site
USD 201,000 - 264,000
Senior Software Engineer, Core Infrastructure
Senior Software Engineer, Core Infrastructure

Harvey • San Francisco (CA)

On-site
USD 200,000 - 250,000
Staff Software Engineer, Site Reliability Engineer (SRE)
Staff Software Engineer, Site Reliability Engineer (SRE)

Harvey • San Francisco (CA)

On-site
USD 238,000 - 290,000
Senior Software Engineer, Production Engineering
Senior Software Engineer, Production Engineering

Harvey • United States

On-site
USD 161,000 - 242,000
Senior Software Engineer, Production Engineering
Senior Software Engineer, Production Engineering

Harvey • New York (NY)

On-site
USD 161,000 - 242,000
Engineering Manager, Production Engineering
Engineering Manager, Production Engineering

Harvey • United States

On-site
USD 260,000 - 340,000
Staff Software Engineer, Production Engineering
Staff Software Engineer, Production Engineering

Harvey • United States

On-site
USD 231,000 - 340,000
Senior Software Engineer, Production Engineering
Senior Software Engineer, Production Engineering

Harvey • San Francisco (CA)

On-site
USD 161,300 - 241,900
Staff Software Engineer, Model Infrastructure
Staff Software Engineer, Model Infrastructure

Neura Market • San Francisco (CA)

On-site
USD 236,000 - 290,000
Staff Software Engineer, Production Engineering Harvey AI New York $231,000 - $340,000/yr
Staff Software Engineer, Production Engineering Harvey AI New York $231,000 - $340,000/yr

Neura Market • New York (NY)

On-site
USD 231,000 - 340,000