Senior Site Reliability Engineer — Remote

Modus Create

Aurora (IL)

Remote

USD 120,000 - 160,000

Full time

13 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Modus Create is seeking an experienced Site Reliability Engineer to build reliable, scalable systems across AWS stack (ALB, ECS/Fargate, Aurora) and Python services. The role is embedded in cross-functional teams to reduce toil and improve uptime, with 24/7 on-call rotations and a focus on proactive reliability rather than firefighting.

You will design and maintain observability, define SLOs and budgets, mentor engineers, and contribute to reliability strategy as the team grows from 4 to 8–12

Qualifications

  • 6+ years hands-on experience in SRE, DevOps, or Platform Engineering at scale.
  • Strong proficiency in AWS (ALB, ECS/Fargate, RDS Aurora, Lambda, IAM).
  • Production Python backend experience; comfortable debugging and optimizing Python services in containers and Lambdas.
  • Experience building or bootstrapping SRE programs from scratch.

Responsibilities

  • Design observability systems (metrics, logging, tracing) and define SLOs, error budgets, and monitoring strategies aligned with business needs.
  • Own 24/7 P0 on-call rotation with 10-minute acknowledgment SLA; validate and escalated AI-generated incident reports to platform teams.
  • Establish reliability standards and SLA targets for backend services.
  • Mentor and onboard additional SRE engineers as team scales to 8-12.
  • Participate in incident response, troubleshooting, root-cause analysis, and postmortems.
  • Implement reliability improvements, capacity planning, and performance optimization.
  • Build self-healing automation and toil-reduction initiatives to minimize manual operational work.
  • Collaborate with teams to integrate reliability and observability into delivery.
  • Support modern workloads including AI applications and services; understand their operational and reliability requirements.
  • Create runbooks, architecture documentation, and troubleshooting guides to enable independence.
  • Review infrastructure changes and contribute to reliability standards and consistency.
  • Actively tune systems for latency, throughput, and resource efficiency based on observability data.

Skills

AWS
Python
Kubernetes
Docker
Terraform
CI/CD
Observability
SLOs
On-call
Linux
Git
Incident response

Tools

Kubernetes
Docker
Terraform
CloudFormation
Pulumi

Job description

Modus Create is seeking an experienced Site Reliability Engineer to build reliable, scalable systems across AWS stack (ALB, ECS/Fargate, Aurora) and Python services. The role is embedded in cross-functional teams to reduce toil and improve uptime, with 24/7 on-call rotations and a focus on proactive reliability rather than firefighting.

You will design and maintain observability, define SLOs and budgets, mentor engineers, and contribute to reliability strategy as the team grows from 4 to 8–12

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer — Remote, AWS & Observability
Senior Site Reliability Engineer — Remote, AWS & Observability

Prove • United States

Hybrid
USD 140,000 - 190,000
Wellbeing reimbursement
401k Match
Parental Leave Policy
+5
Site Reliability Engineer Engineer
Site Reliability Engineer Engineer

Modus Create • Aurora (IL)

Remote
USD 120,000 - 160,000
Remote Site Reliability Engineer - AWS & CI/CD
Remote Site Reliability Engineer - AWS & CI/CD

KnowBe4 • United States

Remote
USD 130,000 - 155,000
Monthly bonuses
Employee referral bonuses
Adoption assistance
+3
Senior Site Reliability Engineer — Remote Production Reliability
Senior Site Reliability Engineer — Remote Production Reliability

Fingerprint • Chicago (IL)

Remote
USD 152,000 - 205,000
Senior Site Reliability Engineer - Remote
Senior Site Reliability Engineer - Remote

EverCommerce • United States

On-site
USD 110,000 - 130,000
Flexible work environment
Robust health and wellness benefits
401(k) with company match
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Axiom Pursuits • San Francisco (CA)

On-site
USD 150,000 - 180,000
Senior Site Reliability Engineer — Flexible Hours, Scale
Senior Site Reliability Engineer — Flexible Hours, Scale

Vertafore • Denver (CO)

On-site
USD 110,000 - 145,000
Senior Site Reliability Engineer — Remote • Flexible PTO
Senior Site Reliability Engineer — Remote • Flexible PTO

ACI Infotech • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 170,000
Health insurance
Dental insurance
Vision insurance
+3
Senior Site Reliability Engineer — Remote Infra & CI/CD
Senior Site Reliability Engineer — Remote Infra & CI/CD

Cross River • United States

Remote
USD 160,000 - 200,000
Remote Senior SRE — Cloud Reliability & Automation
Remote Senior SRE — Cloud Reliability & Automation

EverCommerce • Denver (CO)

Hybrid
USD 110,000 - 130,000
Flexible work environment
Health and wellness benefits
401(k) with company match
+2