Senior Cloud SRE — Serverless Ops & Reliability Lead

Linuxconfig

Northern (KY)

Hybrid

USD 120,000 - 180,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Linuxconfig is seeking a Senior Site Reliability Engineer to join our cloud engineering team. This operations-focused role owns day-to-day administration of AWS accounts and databases, maintaining backup posture across data stores and providing production monitoring and debugging for a serverless platform.

You will lead incident response, manage on-call rotation, and continuously refine infrastructure as code (SST/Pulumi/Terraform) to keep deployments safe, scalable, and cost-visible, while

Qualifications

  • Bachelor's degree and 4–6 years of related experience or equivalent work experience.
  • 5+ years of DevOps, site reliability, or platform operations with significant responsibility for production systems.
  • 3+ years hands-on experience with AWS, emphasizing serverless services (Lambda, SQS, EventBridge, CloudWatch, S3).
  • Strong database administration experience: PostgreSQL operations, backup/recovery, and query performance.
  • Proficiency in scripting languages such as TypeScript, Python, and bash for production automation and tooling.
  • Strong understanding of Linux, DNS, TLS, Docker, GitHub Actions, and infrastructure as code (SST, Pulumi, or Terraform).
  • Experience with production monitoring, incident response, and on-call ownership.

Responsibilities

  • Own day-to-day administration across AWS services, accounts, and access, as well as database administration across PostgreSQL and other data stores.
  • Own backup posture across databases, S3 buckets, and queues; verify restores regularly and maintain a tested disaster recovery plan.
  • Proactively monitor production — CloudWatch dashboards, metric alarms, log-based metrics, and Slack alerting.
  • Lead production debugging and incident response: build and maintain runbooks, participate in the on-call rotation, and resolve queue and dead-letter-queue failures.
  • Continuously refine infrastructure to ensure it is easily deployable and scalable: keep infrastructure as code accurate, retire unused infrastructure, and keep cost visible.
  • Share knowledge of production operations with the team, fostering a culture of learning and growth.

Skills

DevOps
Site Reliability
AWS
Monitoring
Incident response
Scripting
Linux
Docker
IaC
On-call

Education

Bachelor's degree

Tools

SST
Pulumi
Terraform
PostgreSQL
S3
CloudWatch
EventBridge
SQS

Job description

Linuxconfig is seeking a Senior Site Reliability Engineer to join our cloud engineering team. This operations-focused role owns day-to-day administration of AWS accounts and databases, maintaining backup posture across data stores and providing production monitoring and debugging for a serverless platform.

You will lead incident response, manage on-call rotation, and continuously refine infrastructure as code (SST/Pulumi/Terraform) to keep deployments safe, scalable, and cost-visible, while

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud SRE: AWS, Serverless & Incident Response
Senior Cloud SRE: AWS, Serverless & Incident Response

Apply • Northern (KY)

Hybrid
USD 120,000 - 150,000
Senior Cloud SRE: AWS Serverless & Reliability Leader
Senior Cloud SRE: AWS Serverless & Reliability Leader

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 120,000 - 170,000
Senior Cloud SRE: AWS Serverless & Reliability
Senior Cloud SRE: AWS Serverless & Reliability

MeridianLink • United States

Remote
USD 140,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MeridianLink • United States

Remote
USD 140,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Linuxconfig • Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 120,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Apply • Northern (KY)

Hybrid
USD 120,000 - 150,000
Lead Site Reliability Engineer: AWS Cloud & Automation
Lead Site Reliability Engineer: AWS Cloud & Automation

Selby Jennings • Wilmington (NC)

On-site
USD 140,000 - 200,000
Remote Senior SRE — Cloud Ownership & Automation
Remote Senior SRE — Cloud Ownership & Automation

Synthesia • United States

On-site
USD 120,000 - 160,000
Senior Cloud SRE & Reliability Engineer
Senior Cloud SRE & Reliability Engineer

Verygoodsecurity • United States

Hybrid
USD 110,000 - 140,000
Flexible work hours
Competitive health benefits
VGS stock options
+1