Senior SRE — Cloud Native Reliability & Observability

Scientific Games, LLC

Alpharetta (GA)

On-site

USD 140,000 - 190,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Scientific Games is seeking a skilled Site Reliability Engineer (SRE) to boost stability, performance, and reliability of our production systems. You will collaborate with development, DevOps, and security teams to ensure production readiness, manage on-call rotations, and improve observability across applications and infrastructure.

Responsibilities include implementing scalable operations on AWS EKS, automating processes, defining SLIs/SLOs/SLAs, and driving post-incident reviews to prevent

Qualifications

  • Bachelor’s degree in computer science or related field, or equivalent work experience.
  • 6+ years as an SRE, DevOps Engineer, or similar role
  • Cloud: Strong experience with AWS (EKS, EC2, S3, Route53, IAM)
  • Kubernetes: 6+ years managing production Kubernetes workloads
  • Monitoring & Observability: Hands-on with New Relic, Graylog, or similar
  • Secrets Management: Experience with HashiCorp Vault or equivalent
  • Automation & CI/CD: Proficiency with GitHub Actions, GitLab CI/CD, Helm and ArgoCD
  • IaC : Hands-on experience with Terraform
  • Scripting: Proficiency in Python, Bash, or equivalent scripting languages
  • Incident Management: Strong debugging, troubleshooting, and root cause analysis skills
  • On-Call Readiness: Willingness to participate in 24x7 on-call rotation

Responsibilities

  • Enhance production observability with New Relic and Graylog; build dashboards and alerts.
  • Ensure production readiness and high availability through capacity planning and fault-tolerance.
  • Automate operational processes and manage Kubernetes workloads on AWS EKS.
  • Participate in on-call rotation and drive post-incident reviews and fixes.
  • Collaborate with DevOps and development teams to improve CI/CD pipelines.
  • Document runbooks and incident playbooks.

Skills

SRE experience
AWS (EKS/EC2/S3)
Kubernetes
Observability
HashiCorp Vault
CI/CD: GitHub Actions
Terraform
Python/Bash scripting
On-call readiness
Incident management

Education

Bachelor’s degree in CS or related field

Tools

New Relic
Graylog
GitHub Actions
GitLab CI/CD
Helm
ArgoCD
Terraform

Job description

Scientific Games is seeking a skilled Site Reliability Engineer (SRE) to boost stability, performance, and reliability of our production systems. You will collaborate with development, DevOps, and security teams to ensure production readiness, manage on-call rotations, and improve observability across applications and infrastructure.

Responsibilities include implementing scalable operations on AWS EKS, automating processes, defining SLIs/SLOs/SLAs, and driving post-incident reviews to prevent

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Senior SRE - Observability & Automation
Senior SRE - Observability & Automation

PlayOn Sports • United States

Hybrid
USD 140,000 - 210,000
Medical insurance plans
Dental, vision, life and disability
Employee Emergency Fund
+4
Senior SRE: Cloud-Native Reliability for Global SaaS
Senior SRE: Cloud-Native Reliability for Global SaaS

Cisco Systems, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 260,000
Medical insurance
Dental insurance
Vision insurance
+5
Sr. Site Reliability Engineer (SRE)
Sr. Site Reliability Engineer (SRE)

Scientific Games, LLC • Alpharetta (GA)

On-site
USD 140,000 - 190,000
Senior SRE: Scale Reliability, Observability & CI/CD
Senior SRE: Scale Reliability, Observability & CI/CD

Breakout Tools • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior SRE — AWS Reliability & Observability (Remote)
Senior SRE — AWS Reliability & Observability (Remote)

Jobgether • United States

Hybrid
USD 120,000 - 160,000
Competitive compensation package
Flexible work arrangements
Professional development opportunities
+2
Senior SRE: Cloud, Observability & Resilience
Senior SRE: Cloud, Observability & Resilience

Supernova Technology™ • Chicago (IL)

On-site
USD 130,000 - 170,000
Senior SRE: Cloud-Native Reliability & Automation
Senior SRE: Cloud-Native Reliability & Automation

Crypto Pro Network • New York (NY)

On-site
USD 150,000 - 180,000
SRE Lead: Observability & Reliability at Scale
SRE Lead: Observability & Reliability at Scale

OpenBet • Kentucky

Hybrid
USD 140,000 - 200,000
Competitive benefits
Global collaboration
Flexible working
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

myBridge Corporation • Austin (TX)

On-site
USD 120,000 - 160,000