Site Reliability Engineer: Production-Scale AI Infra

Better Tomorrow Ventures

New York (NY)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Health & Wellness benefits
Unlimited PTO
Office meals stipend
401(k) retirement plan
Team events
Parental leave

Job summary

Basis seeks a Site Reliability Engineer to ensure the reliability, scalability, and performance of our AI-powered accounting platform. You will join a high-leverage infrastructure team that owns systems keeping Basis fast, secure, and always on as we scale.

This is a hands-on role for an engineer who loves building robust systems, reducing complexity through automation, and taking real ownership. Strong software engineering skills and willingness to wear multiple hats are essential in our

Qualifications

  • 5+ years of experience building and operating production infrastructure at scale.
  • Strong software engineering fundamentals and proficiency in at least one programming language.
  • Deep understanding of cloud infrastructure, networking, databases, and security principles.
  • Experience with CI/CD systems, containerization, and modern infrastructure automation.

Responsibilities

  • Architect, build, and operate reliable, scalable, and secure infrastructure for our production systems.
  • Own cloud infrastructure across compute, storage, and networking, optimizing for availability, performance, and cost efficiency.
  • Design and maintain CI/CD pipelines, infrastructure-as-code, and automation to improve developer velocity and system reliability.
  • Lead incident response efforts, including on-call rotations, incident coordination, postmortems, and root cause analyses.
  • Partner closely with product and engineering teams to balance reliability, performance, cost, and speed of development.
  • Automate operational workflows such as capacity planning, safe rollouts, graceful degradation, and data access controls.
  • Provide technical leadership and mentorship, helping shape the culture and standards of the infrastructure team.

Skills

Production infrastructure
Cloud infrastructure
CI/CD
Containerization
End-to-end ownership

Tools

Terraform
CloudFormation
OpenTelemetry
Prometheus/Grafana
PagerDuty
SLOs
Neon
Modal

Job description

Basis seeks a Site Reliability Engineer to ensure the reliability, scalability, and performance of our AI-powered accounting platform. You will join a high-leverage infrastructure team that owns systems keeping Basis fast, secure, and always on as we scale.

This is a hands-on role for an engineer who loves building robust systems, reducing complexity through automation, and taking real ownership. Strong software engineering skills and willingness to wear multiple hats are essential in our

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – AI Platform Infra (Onsite NYC)
Site Reliability Engineer – AI Platform Infra (Onsite NYC)

getbasis.ai • New York (NY)

On-site
USD 140,000 - 190,000
Health & Wellness benefits
Time off — unlimited PTO + holidays
In-Office perks — meals, kitchen, desk
Site Reliability Engineer — Scale an AI‑Powered SaaS Platform
Site Reliability Engineer — Scale an AI‑Powered SaaS Platform

Instrumental Inc. • Palo Alto (CA)

On-site
USD 140,000 - 165,000
Health insurance
Vision insurance
Dental plan
+2
Senior Site Reliability Engineer — AI Platform Scale
Senior Site Reliability Engineer — AI Platform Scale

Future Secure AI • Austin (TX)

On-site
USD 140,000 - 190,000
Site Reliability Engineer — ML Infra, Scale & Equity
Site Reliability Engineer — ML Infra, Scale & Equity

Baseten • New York (NY)

On-site
USD 165,000 - 330,000
Site Reliability Engineer — Scale & Resilience for AI Ops
Site Reliability Engineer — Scale & Resilience for AI Ops

HappyRobot • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior SRE – AI Infrastructure Reliability Leader
Senior SRE – AI Infrastructure Reliability Leader

Nscale • San Francisco (CA), Seattle (WA), Houston (TX)

On-site
USD 170,000 - 265,000
Equity
Ownership from start
Flexible schedule
Senior Site Reliability Engineer - AI-Driven Infra
Senior Site Reliability Engineer - AI-Driven Infra

TENEX.AI • Sarasota (FL)

On-site
USD 150,000 - 230,000
Senior Backend Engineer - AI-Driven Reliability Platform
Senior Backend Engineer - AI-Driven Reliability Platform

Affirm • Portland (OR)

On-site
USD 173,000 - 233,000
Health care coverage
Flexible Spending Wallets
Time off
+1
Platform Engineer - Build AI-Scale Foundations
Platform Engineer - Build AI-Scale Foundations

Scale AI, Inc. • New York (NY)

On-site
USD 216,000 - 270,000
Health, dental & vision
Retirement benefits
Learning & development stipend
+2
Site Reliability Engineer for AI Factory Infra (On-Site)
Site Reliability Engineer for AI Factory Infra (On-Site)

1872 Ai • Cincinnati (OH)

On-site
USD 140,000 - 190,000
Significant equity
Unlimited PTO
Paid parental leave
+1