Senior Infrastructure Engineer

Manufact (YC S25)

San Francisco (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Manufact (YC S25), based in San Francisco, is seeking a cloud engineer to build and operate their cloud platform hosting customer MCP servers. You will set up observability systems, ensure security, and maintain system reliability as user load increases.

The ideal candidate has extensive experience with cloud systems, Kubernetes, and observability tools. Join a fast-growing startup and take ownership of critical infrastructure design and implementation.

Qualifications

  • Experience building and operating cloud systems end to end.
  • Knowledge of observability tools for monitoring cloud health.
  • Familiarity with multi-tenant systems and infrastructure ownership.

Responsibilities

  • Build the cloud platform that hosts customer MCP servers.
  • Set up monitoring and observability across all MCP servers.
  • Create tooling for customers to deploy Manufact in private clouds.

Skills

Cloud system operations
Observability stack experience
Kubernetes
Terraform
Containerized systems

Tools

ClickHouse
Prometheus
OpenTelemetry
Grafana

Job description

What you'll do
  • Build the cloud platform that hosts customer MCP servers with millions of tool calls.
  • Set up monitoring, observability, and dashboards across every MCP server on Manufact Cloud and the full stream of events and logs we ingest, so problems surface before users feel them.
  • Build alerting and error handling that flags threshold breaches and failures in real time.
  • Make sure the platform scales as we add thousands of users and capture a growing share of global AI tool calls.
  • Create the tooling and technical blueprints that let customers deploy Manufact into their own private clouds (AWS, Azure, GCP) on their own terms.
  • Build the security boundaries that keep each customer's data private and isolated within their own environment.
  • Own where our code lives and how it ships: code hosting and the release pipeline.
  • Keep the system fast and reliable as companies add more MCP servers.
Requirements
  • You have built and operated cloud systems end to end and know them inside out.
  • Experience standing up the observability stack that keeps a cloud healthy: logs, metrics, traces, dashboards, and alerting at scale.
  • Kubernetes or Terraform, and containerized multi-tenant systems.
  • Comfortable being the first person to own infrastructure at a fast-growing startup.
  • Based in or willing to relocate to San Francisco.
Nice to have
  • Hands‑on with observability tooling such as ClickHouse, Prometheus, OpenTelemetry, Grafana.
  • Experience running a platform that others deploy onto (PaaS‑flavored).
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Platform Engineer - Scale & Observability
Senior Cloud Platform Engineer - Scale & Observability

Manufact (YC S25) • San Francisco (CA)

On-site
USD 120,000 - 150,000
DevOps & Cloud Infrastructure Engineer
DevOps & Cloud Infrastructure Engineer

Methodic • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Software Engineer – MCP & AI Platform
Senior Software Engineer – MCP & AI Platform

InApp • United States

Hybrid
USD 140,000 - 190,000
Engineering Manager, Cloud Infrastructure
Engineering Manager, Cloud Infrastructure

Jobtailor • Washington

On-site
USD 170,000 - 210,000
Senior Cloud Engineer
Senior Cloud Engineer

Insight Global • Atlanta (GA)

On-site
USD 140,000 - 190,000
Senior/Staff Engineer
Senior/Staff Engineer

Pagos Consultants • New York (NY)

Hybrid
USD 250,000 - 350,000
Senior AI Platform Security Engineer - Contract to Hire San Francisco, CA
Senior AI Platform Security Engineer - Contract to Hire San Francisco, CA

FrontApp Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Platform Engineer
Platform Engineer

Synergy • Chicago (IL)

On-site
USD 100,000 - 150,000
Member of Technical Staff [Platform]
Member of Technical Staff [Platform]

NeoCognition Inc. • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Senior DevOps Engineer
Senior DevOps Engineer

Strativ Group • San Francisco (CA)

Hybrid
USD 225,000 - 275,000
Flexible PTO
HSA contribution
Parental leave
+1