Remote Senior Site Reliability Engineer (US or Canada)

MAP SSG

United States

Remote

USD 160,000 - 210,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity compensation

Job summary

MAP SSG is seeking a Senior Site Reliability Engineer to join a 100% remote team in the US or Canada. You will own CI/CD end-to-end, build platform capabilities, and help extend an AI-agent harness for autonomous development workflows.

The role emphasizes DevEx, platform reliability, and strong collaboration across distributed teams. Ideal candidates have 6+ years in SRE or related fields, hands-on experience with TS/Node.js, Python, Terraform, Kubernetes/Helm, and robust observability

Qualifications

  • 6+ years of SRE/platform engineering or related roles.
  • Experience owning CI/CD platforms end to end.
  • Experience with CI/CD concepts: caching, architecture, build/test optimization.
  • Strong software fundamentals for coding interviews.
  • Familiarity with TS/Node.js, Python, Terraform, Kubernetes/Helm.

Responsibilities

  • Own and evolve CI/CD systems end to end, including architecture and self-service.
  • Build tooling to reduce toil and accelerate safe shipping.
  • Contribute to an AI-agent harness with CI integration and sandboxing.
  • Improve reliability for a high-volume real-time AI platform.
  • Maintain infrastructure using Terraform, Kubernetes, and Helm.
  • Improve observability across logs, metrics, traces, and alerts.
  • Collaborate with engineers to troubleshoot production issues.
  • Set patterns and standards across the engineering org.
  • Participate in on-call rotation; after-hours paging is uncommon.

Skills

CI/CD ownership
DevEx
SRE experience
Remote work experience
TypeScript/Node.js
Python
Terraform
Kubernetes/Helm
Observability
Cloud platforms (GCP)

Tools

Terraform
Kubernetes
Helm
Datadog
Prometheus
Grafana
GitLab CI
GCP

Job description

Remote Senior Site Reliability Engineer (US or Canada)
  • US, Remote
Senior Site Reliability Engineer

Remote — United States or Canada
$160-210K base + equity

About the Company

We’re partnering with a well-funded Series B AI company building an enterprise platform at the intersection of voice AI, large language models, and real-time customer interactions. The company has raised more than $100M, supports large enterprise customers, and has been running AI agents in production for years—not simply adding AI to an existing product.

The engineering organization is intentionally lean and senior. This is a 100% remote company across the U.S. and Canada, with engineers owning multiple initiatives and operating with significant autonomy.

The Role

This is not a traditional infrastructure-only SRE position.

You’ll join a small, senior platform/SRE team responsible for the systems that enable the broader engineering organization to build, test, deploy, observe, and operate an AI-native production platform.

The team owns areas including:

  • CI/CD and developer experience
  • Platform engineering and reliability
  • Developer tooling and paved paths
  • Observability and incident management
  • Kubernetes and cloud infrastructure
  • Cloud cost visibility and optimization
  • AI-agent harness engineering
  • Sandboxes, guardrails, validation, and agent‑first development workflows

The hiring manager is particularly interested in engineers with strong DevEx, CI/CD, platform engineering, or SRE backgrounds rather than candidates whose experience is primarily compute, networking, data centers, or traditional infrastructure administration.

What You’ll Do
  • Own and evolve CI/CD systems end to end, including architecture, caching, build/test performance, deployment workflows, and developer self‑service.
  • Build tooling and platform capabilities that reduce engineering toil and help developers ship safely and quickly.
  • Help extend an internal AI-agent harness used to support autonomous development workflows, including CI integration, sandboxing, guardrails, and validation.
  • Improve reliability and operability for a high‑volume, real‑time AI platform.
  • Build and maintain infrastructure using Terraform, Kubernetes, and Helm.
  • Improve observability across logs, metrics, traces, monitoring, and alerting.
  • Partner directly with software engineers to troubleshoot production and development issues, including making changes within application code where needed.
  • Help establish patterns and technical standards across the engineering organization.
  • Participate in an on‑call rotation focused on base infrastructure; after‑hours pages are uncommon.
What We’re Looking For
  • 6+ years of experience in SRE, platform engineering, DevEx, developer infrastructure, or related software‑development enablement roles.
  • Strong experience owning CI/CD platforms end to end, rather than simply maintaining existing pipelines.
  • Experience with CI/CD concepts such as caching, architecture, build/test optimization, deployment ergonomics, and developer self‑service.
  • Strong enough software engineering fundamentals to succeed in a coding‑oriented technical interview.
  • Working familiarity with:
    • TypeScript / Node.js
    • Python
    • Terraform
    • Kubernetes / Helm
  • Practical observability experience across logs, metrics, tracing, monitoring, alerting, and incident management.
  • Experience working successfully on distributed or fully remote engineering teams.
  • Understanding of how modern LLMs and AI development tools work.
  • Hands‑on use of tools such as Claude, Cursor, Copilot, or similar—and the judgment to know when AI‑generated output should not be trusted without deeper evaluation.

The environment currently includes TypeScript, Node.js, Python, Kubernetes, Helm, GCP, GitLab CI, Terraform, Datadog, Prometheus, and Grafana.

Especially Interesting Backgrounds

We’d be particularly interested in engineers who have worked on:

  • Developer platforms or internal developer infrastructure
  • CI/CD infrastructure at meaningful scale
  • Developer productivity / DevEx
  • Cloud‑native SaaS platforms
  • AI or ML infrastructure
  • Platforms for autonomous or semi‑autonomous AI agents
  • Large‑scale GCP environments
  • Real‑time communications, telephony, SIP, or FreeSWITCH

Experience in voice AI or conversational AI is helpful, but not required.

What This Role Is Not

This is probably not the right fit if your background is primarily:

  • Data center or on‑prem infrastructure
  • Cloud compute administration without meaningful DevEx or software engineering ownership
  • Infrastructure operations with little hands‑on coding
  • Maintaining CI/CD pipelines without having designed or owned the platform behind them
Team & Culture

You’ll join a small, senior engineering organization with staff‑and principal‑level engineers and substantial individual ownership. There are no junior engineers on the team.

The culture rewards people who:

  • Take ownership and drive projects forward.
  • Are comfortable operating across multiple technical domains.
  • Will tackle both large architectural problems and unglamorous operational work.
  • Move quickly, experiment, learn from mistakes, and iterate.
  • Can critically evaluate their own decisions and respond well to feedback.
Compensation & Location

Base salary: approximately $160,000-210,000 USD, depending on level and location, plus competitive equity. Compensation may vary for Canadian employees.

Location: Fully remote within the United States or Canada, working primarily across U.S. time zones.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE / Cloud / Kubernetes / Terraform / 100% Remote
Senior SRE / Cloud / Kubernetes / Terraform / 100% Remote

Motion Recruitment • United States

Remote
USD 140,000 - 170,000
Medical, dental, and vision
Equity / Stock Options
Remote equipment stipend
+3
Senior SRE / Cloud / Kubernetes / Terraform / 100% Remote
Senior SRE / Cloud / Kubernetes / Terraform / 100% Remote

Motion Recruitment • Mount Laurel Township (NJ)

Remote
USD 130,000 - 180,000
Medical, dental, and vision benefits
Equity / Stock Options
Remote equipment stipend
+3
Senior SRE for AI-Native Platform — Remote US/Canada
Senior SRE for AI-Native Platform — Remote US/Canada

MAP SSG • United States

Remote
USD 160,000 - 210,000
Equity compensation
Senior DevOps Engineer
Senior DevOps Engineer

Talener • Kansas

Remote
USD 160,000 - 180,000
Comprehensive benefits
Remote-first / Remote-friendly
Senior SRE
Senior SRE

banyansoftware • United States

Remote
USD 130,000 - 165,000
Remote role (US/Canada)
Senior Forward Deployed Engineer (DevOps/SRE)
Senior Forward Deployed Engineer (DevOps/SRE)

LeoForce • Pleasanton (CA)

On-site
USD 300,000 - 350,000
Medical benefits
401(k) plan
Free meals and snacks
+2
Site Reliability Engineer Engineer
Site Reliability Engineer Engineer

Modus Create • Aurora (IL)

Remote
USD 120,000 - 160,000
Senior Software Engineer
Senior Software Engineer

ProNexus • New York (NY)

On-site
USD 130,000 - 160,000
Fully remote
Competitive equity
Team Leader, SRE
Team Leader, SRE

Remote • United States

Remote
USD 75,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hard Rock Digital • United States

Hybrid
USD 150,000 - 210,000