Remote Senior Site Reliability Engineer — AI‑Driven Infra

Precisely

Atlanta (GA)

On-site

USD 150,000 - 190,000

Full time

6 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Precisely is seeking a Senior Site Reliability Engineer to own reliability for CCX, RCX, and HMS across on‑prem, private cloud, and AWS. You will drive automation, observability, and incident response, partnering with engineering teams to raise production readiness.

You will mentor peers, define SLOs, and lead operational readiness reviews, change reviews, and disaster recovery planning in a fast‑moving environment.

Qualifications

  • Bachelor's degree in Computer Science, Information Systems, Engineering, or equivalent Prac tical experience.
  • 5+ years of systems or infrastructure engineering experience in an enterprise production environment.
  • High-level competency in one or more complex infrastructure domains (cloud platform engineering, IaC automation, observability, or on-prem virtualization).
  • Advanced proficiency with Linux (RHEL/Oracle Linux) across on-prem and cloud environments.
  • Proficiency with Terraform for infrastructure-as-code; experience with Ansible beyond basic usage.
  • Experience deploying and managing workloads in AWS (EC2, ECS, S3, VPC, IAM, CloudWatch, Auto Scaling).
  • Proficiency with Python or Bash for automation.
  • Experience designing monitoring and alerting (Datadog preferred).
  • Ability to define SLOs and lead Operational Readiness Reviews.

Responsibilities

  • Define and maintain reliability standards across CCX, RCX, and HMS.
  • Define alerting, logging, and tracing standards; partner with engineering teams.
  • Lead complex reliability work across on-prem, private cloud, and AWS.
  • Build and maintain Infrastructure-as-Code, deployment automation, monitoring tools.
  • Embed reliability and operability considerations into service design and delivery.
  • Lead ORRs and validate production readiness and DR readiness for changes.
  • Coordinate release triage, deployment, and on-call duties.
  • Lead incident response as incident commander when needed; communicate status clearly.
  • Draft root-cause analyses and preventive automation to reduce toil and MTTR.
  • Maintain runbooks, reliability backlogs, and knowledge resources.

Skills

Linux proficiency
Terraform
Ansible
AWS
Python scripting
Datadog / observability
CI/CD pipelines
SRE fundamentals

Education

Bachelor's degree in CS or related field

Tools

Terraform
Ansible
Datadog
AWS CLI

Job description

Precisely is seeking a Senior Site Reliability Engineer to own reliability for CCX, RCX, and HMS across on‑prem, private cloud, and AWS. You will drive automation, observability, and incident response, partnering with engineering teams to raise production readiness.

You will mentor peers, define SLOs, and lead operational readiness reviews, change reviews, and disaster recovery planning in a fast‑moving environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: AI-Driven Reliability & Cloud Automation
Senior SRE: AI-Driven Reliability & Cloud Automation

NDEAVOUR CONSULTING • United States

Hybrid
USD 120,000 - 150,000
Remote Office
Parking Space
Fun Office Space
+7
Senior SRE: AI-Driven Reliability & Automation (Hybrid)
Senior SRE: AI-Driven Reliability & Automation (Hybrid)

Namely • United States

Hybrid
USD 120,000 - 150,000
Remote Senior Site Reliability Engineer — Reliability Lead
Remote Senior Site Reliability Engineer — Reliability Lead

Priority Technology Holdings, Inc. • Alpharetta (GA)

On-site
USD 129,000 - 161,000
401(k) match
Employee Stock Purchase Program (ESPP)
Medical, dental, and vision coverage
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SDI International • Chicago (IL)

Hybrid
USD 130,000 - 180,000
Senior SRE - Multi-Cloud Reliability & AI-Driven Ops
Senior SRE - Multi-Cloud Reliability & AI-Driven Ops

Satsuma AI, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 140,000 - 210,000
Unlimited PTO
401(K)
Healthcare Stipend
+1
Senior SRE — Scale, Automation & Uptime
Senior SRE — Scale, Automation & Uptime

Hirebridge • Northern (KY)

Hybrid
USD 110,000 - 145,000
Bonus
Senior Site Reliability Engineer — AI Platform Scale
Senior Site Reliability Engineer — AI Platform Scale

Future Secure AI • Austin (TX)

On-site
USD 140,000 - 190,000
Remote Senior SRE: Cloud Reliability, CI/CD & AI-Driven
Remote Senior SRE: Cloud Reliability, CI/CD & AI-Driven

IDEXX • Boston (MA)

On-site
USD 100,000 - 125,000
Health benefits
401k matching
Pet Insurance
+1
Remote SRE Manager: Lead AI-Driven Reliability & Cloud Ops
Remote SRE Manager: Lead AI-Driven Reliability & Cloud Ops

Arcoro Holdings Corp • Phoenix (AZ), Northern (KY)

Hybrid
USD 200,000 - 220,000
Remote Work
401(k) with Company match
Flexible PTO and Company-paid holidays
Senior SRE – AI Infrastructure Reliability Leader
Senior SRE – AI Infrastructure Reliability Leader

Nscale • San Francisco (CA), Seattle (WA), Houston (TX)

On-site
USD 170,000 - 265,000
Equity
Ownership from start
Flexible schedule