Infrastructure Engineer

Reducto, Inc.

San Francisco (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Free lunch
Reimbursed transportation
Health insurance
Health and wellness budget
Parental leave

Job summary

Reducto, Inc. is seeking an Infrastructure Engineer to design, build, and maintain scalable infrastructure for AI and ML workloads. The role involves automating cloud infrastructure and implementing robust monitoring systems to ensure reliability.

With a requirement of 5+ years of experience and a strong focus on quality solutions, this in-person position is based in San Francisco and offers several benefits.

Qualifications

  • 5+ years of experience in production-grade infrastructure.
  • Experience with AI/ML workloads and high-throughput systems.
  • Strong ability to implement monitoring and observability systems.

Responsibilities

  • Design, build, and maintain scalable infrastructure for AI/ML.
  • Implement monitoring and alerting systems for system health.
  • Debug and optimize infrastructure for rapid deployment.

Skills

Linux systems
Cloud platforms
Python
Kubernetes
Networking
Storage technologies

Job description

Reducto is the agentic document platform for leading AI teams who demand enterprise performance at scale. We provide a comprehensive toolkit for working with documents the way a human would, combining custom in‑house and leading frontier models to power efficient and accurate document workflows.

We’ve grown rapidly, increasing revenue 8x year over year and partnering with hundreds of companies, from leading AI teams like Harvey, Vanta, and Scale, to enterprise customers across FAANG and top trading firms.

Reducto has raised over $100M from world‑class investors including a16z, Benchmark, and First Round Capital.

The Opportunity

As an Infrastructure Engineer at Reducto, you will influence every aspect of our infrastructure from the ground up. You will architect and scale resilient systems for AI and ML workloads, automate cloud infrastructure, and implement monitoring and incident response practices that set the standard for reliability. This role requires technical leadership, hands‑on systems engineering, and strong collaboration with our founders and product teams as we build a company around reliability, rapid iteration, and high‑impact product delivery.

The core work will include:
  • Designing, building, and maintaining highly available, scalable infrastructure to support intensive AI/ML workloads and real‑time model deployments.
  • Implementing robust monitoring, alerting, and observability systems to ensure system health, performance, and uptime across cloud and on‑prem environments.
  • Debugging, optimizing, and automating infrastructure for fast iteration and rapid deployment cycles, focusing on both reliability and developer velocity.
  • Proactively identifying, investigating, and resolving incidents to minimize downtime and maintain world‑class service levels for enterprise customers.
  • Collaborating closely with engineers, ML specialists, and founders to shape product, infrastructure, and security strategies.
We would love to meet you if you:
  • Are your own worst critic—have an extremely high bar for quality and always aim for robust solutions rather than quick fixes.
  • Have 5+ years of hands‑on experience in building or supporting production‑grade infrastructure and reliability processes for high‑throughput systems.
  • Are comfortable with Python or similar languages, and exceptional at working across cloud platforms, container orchestration (e.g., Kubernetes), networking, and storage technologies.
  • Build your own tools on the fly to diagnose, experiment, and address reliability problems—whether it's an internal dashboard or an automated remediation workflow.
  • Bring a quantitative, hands‑on approach to system operations, automation, and continuous improvement.
  • Have prior experience founding a company or building products/infrastructure in early‑stage, high‑growth environments.
  • Are excited about automating incident management processes with LLMs/AI.
  • Are driven, ambitious, and deeply care about both technical excellence and collaborative problem‑solving.
  • Keep up with the latest trends in cloud, observability, and SRE best practices.
  • Are passionate about open‑source and have contributed tools or automation to reliability communities.
  • Have built or optimized monitoring, incident response, or high‑performance computing systems for demanding AI/ML, fintech, or enterprise clients.

This is an in person role at our office in SF. We’re an early stage company which means that the role requires working hard and moving quickly. Please only apply if that excites you.

Benefits at Reducto
  • Lunch: Receive a free lunch to eat with your teammates daily at the office
  • Reimbursed Transportation: Provide us with your receipts and we’ll take care of the costs
  • Insurance: Generous health insurance covering medical, dental, and vision.
  • Health and Wellness Budget: We provide up to $150/mo reimbursement for health and wellness spending, such as gym memberships, fitness classes, or similar.
  • Parental Leave: Work with us to build a leave schedule that works for you and your family

Reducto is an Equal Opportunity Employer committed to diversity and inclusion in the workplace. All qualified applicants will receive consideration for employment without regard to sex, race, color, age, national origin, religion, physical and mental disability, genetic information, marital status, sexual orientation, gender identity/assignment, citizenship, pregnancy or maternity, protected veteran status, or any other status prohibited by applicable national, federal, state or local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Infrastructure Engineer
Infrastructure Engineer

Reducto • San Francisco (CA)

On-site
USD 120,000 - 160,000
Unlimited PTO
Free daily lunch
Reimbursed transportation
+3
Forward Deployed Engineer, Infrastructure Specialist
Forward Deployed Engineer, Infrastructure Specialist

Reducto, Inc. • San Francisco (CA)

On-site
USD 120,000 - 150,000
Free lunch
Reimbursed transportation
Generous health insurance
+2
Forward Deployed Engineer
Forward Deployed Engineer

Reducto • San Francisco (CA)

On-site
USD 120,000 - 150,000
Unlimited PTO
Free lunch daily
Reimbursed transportation
+3
Forward Deployed Engineer, Infrastructure Specialist
Forward Deployed Engineer, Infrastructure Specialist

Reducto • San Francisco (CA)

On-site
USD 120,000 - 150,000
Unlimited PTO
Free daily lunch
Reimbursed transportation
+3
Solutions Engineer
Solutions Engineer

Reducto, Inc. • San Francisco (CA)

On-site
USD 140,000 - 210,000
Free daily lunch
Reimbursed transportation costs
Generous health insurance
+2
Lead Software Engineer, Platform
Lead Software Engineer, Platform

Reducto • San Francisco (CA)

On-site
USD 150,000 - 180,000
Unlimited PTO
Free lunch
Reimbursed transportation
+3
Lead Software Engineer, Platform
Lead Software Engineer, Platform

Reducto, Inc. • San Francisco (CA)

On-site
USD 180,000 - 240,000
Unlimited PTO
Lunch
Reimbursed Transportation
+3
Machine Learning Infra Engineer
Machine Learning Infra Engineer

Reducto • San Francisco (CA)

On-site
USD 120,000 - 160,000
Unlimited PTO
Free lunch
Reimbursed transportation
+3
Solutions Engineer
Solutions Engineer

Reducto • San Francisco (CA)

On-site
USD 120,000 - 180,000
Unlimited PTO
Free daily lunch
Reimbursed transportation
+3
Machine Learning Infra Engineer
Machine Learning Infra Engineer

Reducto, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Free lunch
Reimbursed transportation costs
Generous health insurance
+2