Infrastructure Software Engineer

mercor

San Francisco (CA)

On-site

USD 180,000 - 280,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Performance bonus
Equity grant
Relocation bonus
Housing stipend
Meals stipend
Equinox membership
Laundry reimbursement
Wellness reimbursement
Health insurance

Job summary

Mercor is hiring an Infrastructure Engineer in the San Francisco/New York area to design and scale systems that power rapid growth. You will build multi-tenant platforms, ensure high availability, and drive observability across production services.

The role focuses on designing scalable architectures, deploying tooling, and owning the reliability program. You will work with Python/Go, AWS, Terraform, PostgreSQL, MongoDB, and Kubernetes, and collaborate with product, research, and operations

Qualifications

  • Focus on how you reason about systems over tool count.
  • Experience operating production systems used by real users.
  • Strong software engineering fundamentals with Python or Go.
  • AWS cloud experience and Terraform IaC.
  • Knowledge of containers and Kubernetes is a plus.

Responsibilities

  • Design and build multi-tenant platforms with isolation and observability.
  • Develop internal tools from deploy tooling to on-call automation.
  • Scale multi-tenant services across model providers with quotas.
  • Operate Temporal, PostgreSQL, MongoDB, and Kubernetes at scale.
  • Define SLOs, reduce alerts, and improve incident handling.
  • Create a developer platform for a fast-shipping engineering org.
  • Define network and identity boundaries across environments.

Skills

Reliability basics
Scalability understanding
Distributed systems
Production systems experience
Python or Go
AWS cloud
Terraform
CI/CD practices
On-call ownership

Tools

PostgreSQL
MongoDB
Temporal
Kubernetes

Job description

About Mercor

Mercor\'s mission is to organize human intelligence to power the AI economy. We\'re a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor\'s APEX benchmark family measures AI\'s real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents.

Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You\’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.

About the Role

As an Infrastructure Engineer at Mercor, you\’ll build and scale the systems that power our rapid growth. You\’ll ensure our infrastructure is highly available, cost-effective, and able to handle explosive traffic and compute demands. You\’ll work closely with engineers across product, research, and operations to design scalable architectures, streamline deployments, and improve observability.

This is a software engineering role. You\'ll spend most of your time designing and writing the services, platforms, and tooling that the rest of engineering builds on, and the rest making sure they hold up in production.

We\'re hiring across Infrastructure: Platform, Developer Productivity, Production Engineering, Storage and Databases. We team-match after the first screen, so apply even if your background leans toward one area.

What You\'ll Work On
  • Design and build the platforms that sit between our engineers and the outside world: multi-tenant services that give dozens of teams isolated capacity, routing, quotas, and observability on demand, owned from architecture review through production.
  • Build the internal platforms and solutions that product and research engineers rely on every day, from deploy tooling to on-call automation, and treat them with the same design and testing rigor as customer-facing code.
  • Scale our multi-tenant services that fronts every model provider we use, so dozens of internal teams and products get isolated quotas, routing, observability, and capacity on demand without filing a ticket.
  • Run Temporal, Postgres, MongoDB, and our Kubernetes fleet at a scale where "it worked last month" isn\'t a guarantee, and design for the next 10x.
  • Own the reliability program: define SLOs that matter, kill noisy alerts, make CI/CD deploys boring, and turn every incident into a durable fix rather than a runbook entry.
  • Build the developer platform for an engineering org that ships dozens of times a day, and where coding agents are now first-class users of our CI, sandboxes, and deploy pipelines.
  • Design network and identity boundaries across production, preprod, and research compute so teams move fast without cross-environment risk.
  • Make the cost picture legible: attribute spend across data centers, and inference providers, and find the architectural changes that bend the curve.
  • Build the tooling that lets the rest of engineering self-serve: our on-call is already AI-triaged and auto-assigned; you\'ll decide what gets automated next.
What We\'re Looking For
  • We care far more about how you reason about systems than which tools you\'ve used. Strong candidates typically have:
  • A deep grasp of reliability and scalability fundamentals: failure modes, backpressure, idempotency, capacity planning, and how distributed systems actually break under load.
  • Experience operating production systems that real users depend on, and the scars to prove it.
  • Strong software engineering fundamentals: you design systems before you build them, write code you\'re proud of in Python or Go, test it, and review others\' code with care. Most of your recent work has been shipping software, not clicking through consoles.
  • Familiarity with cloud infrastructure (we\'re on AWS) and infrastructure-as-code (we use Terraform). You don\'t need to be an expert; you need to be curious and fast to pick it up.
  • Working knowledge of containers and how they\'re deployed. Kubernetes experience is a plus, not a gate.
  • High ownership: you see a gap, you write the proposal, you ship it, you carry the pager for it.
You Might Be a Great Fit If
  • You\’re a backend engineer with strong infra fundamentals who keeps getting pulled toward the platform layer because that\’s where the hardest problems are.
  • You\’ve been the person who understood why the system fell over when nobody else did.
  • You\’ve built internal platforms or tooling that other engineers loved using.
  • You\’ve operated databases, message queues, or workflow engines at meaningful scale.
  • You\’ve come from backend or DevOps and want to own infrastructure end to end.
How We Work
  • Small, senior team with direct access to the head of infra and to engineering leadership.
  • In-person five days a week in SF, or NYC. Infra is a team sport for us.
  • We use AI coding agents heavily and expect you to as well; the interesting work is deciding what they should do, not typing.
  • Weekly on-call rotation with structured handoffs; on-call load is a metric we actively drive down.
Benefits
  • Bi-annual performance bonus structure
  • Generous equity grant vested over 4 years
  • Up to $15k Relocation bonus
  • $10K housing bonus (if you live within 0.5 miles of our office)
  • $1.5K monthly stipend for meals
  • Free Equinox membership
  • $200 monthly laundry reimbursement
  • $200 monthly personal wellness reimbursement
  • Health, Dental, Vision insurance
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer, Frontier Data Products
Software Engineer, Frontier Data Products

Mercor • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Bi-annual bonus
Generous equity
Relocation bonus
+3
Software Engineer, Backend
Software Engineer, Backend

Mercor, Inc. • San Francisco (CA)

On-site
USD 190,000 - 250,000
Bi-annual bonus
Equity grant
Relocation bonus
+6
Member of Technical Staff, Tech Lead Applied AI Backend
Member of Technical Staff, Tech Lead Applied AI Backend

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Bi-annual bonus
Equity grant
Relocation bonus
+5
Software Engineer, Enterprise Applied AI
Software Engineer, Enterprise Applied AI

Mercor • New York (NY)

On-site
USD 140,000 - 210,000
Bi-annual performance bonus
Equity grant
Relocation bonus
+6
Member of Technical Staff, Applied AI Backend
Member of Technical Staff, Applied AI Backend

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Relocation bonus
Housing stipend
Meals stipend
+2
Software Engineer, Frontier Data Products
Software Engineer, Frontier Data Products

Alumni Ventures • San Francisco (CA)

On-site
USD 120,000 - 160,000
Bi-annual performance bonus
Generous equity grant
Relocation bonus up to $15k
+2
Software Engineer, Enterprise Applied AI
Software Engineer, Enterprise Applied AI

Mercor • San Francisco (CA)

On-site
USD 180,000 - 260,000
Performance bonus
Equity grant
Relocation bonus
+6
Member of Technical Staff, Backend
Member of Technical Staff, Backend

Mercor • San Francisco (CA)

On-site
USD 180,000 - 280,000
Generous equity grant
relocation bonus
housing bonus
+3
Fullstack Software Engineer, Agent Platform
Fullstack Software Engineer, Agent Platform

Mercor • San Francisco (CA)

On-site
USD 140,000 - 200,000
Bi-annual bonus
Equity grant
Relocation bonus
+8
Software Engineer, Platform
Software Engineer, Platform

Mercor • San Francisco (CA), New York (NY)

On-site
USD 120,000 - 160,000
Bi-annual performance bonus structure
Generous equity grant vested over 4 years
Up to $15k Relocation bonus
+6