Senior Site Reliability Engineer

Apex Systems

Louisville (KY)

Hybrid

USD 120,000 - 180,000

Full time

23 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Apex Systems is seeking an experienced Senior Site Reliability Engineer to own reliability and automation across cloud and edge-based restaurant technology deployments in Louisville, KY. You will lead SRE practices, design scalable infrastructure, and collaborate with security and DevOps teams to raise operational excellence.

The role emphasizes hands-on engineering, incident reviews, and governance of cloud architectures with a focus on edge systems and MDM platforms.

Qualifications

  • 6+ years in IT infrastructure, with 3+ years focused on site reliability engineering, cloud architecture, or platform engineering.
  • Hands-on experience with AWS services (e.g., VPC, EC2, ECS, EKS, Fargate, Lambda, IAM, RDS, S3).
  • Working experience with Microsoft Azure.
  • Strong expertise in Kubernetes and container orchestration, including edge or distributed deployments.
  • Experience with GitLab CI/CD and CI/CD pipeline design.
  • Solid experience with Terraform for infrastructure provisioning.
  • Experience with enterprise networking, including switch configuration, VLANs, and network troubleshooting.
  • Familiarity with Mobile Device Management (MDM) platforms.
  • Experience with automation and scripting (Python, Bash, Go, or equivalent).
  • Proven ability to build internal tooling and API integrations.
  • Experience defining and operating against SLOs, SLIs, and error budgets.
  • Comfortable working in Linux command-line environments.

Responsibilities

  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability with velocity.
  • Design, build, and maintain monitoring, alerting, and observability platforms that provide meaningful signals.
  • Lead blameless post-incident reviews, drive root cause analysis, and ensure permanent corrective actions are implemented.
  • Design, implement, and support secure, scalable, and highly available cloud solutions primarily within AWS, with working knowledge of Azure.
  • Architect and manage containerized workloads using Kubernetes, including both cloud and edge-based deployments.
  • Develop and maintain Infrastructure as Code (IaC) using Terraform.
  • Build and optimize CI/CD pipelines using GitLab CI/CD and modern DevOps practices.
  • Develop automation solutions, internal tools, and middleware integrations to eliminate repetitive operational work.
  • Design and support Kubernetes-based edge systems deployed in restaurant locations.
  • Manage and optimize Mobile Device Management (MDM) platforms.
  • Establish and enforce cloud governance, security policies, and architectural standards.
  • Provide technical leadership and mentorship to engineering teams.

Skills

Kubernetes
Cloud architecture
SRE practices
Observability
Edge deployments
Linux
Automation
Security

Tools

Terraform
GitLab CI/CD
Python
Bash
Go

Job description

Overview

An experienced and pragmatic Senior Site Reliability Engineer is sought to own the reliability, design, implementation, and continuous improvement of the infrastructure that powers restaurant technology. This includes cloud platforms, CI/CD pipelines, Kubernetes-based edge systems deployed in restaurants, networks, MDM platforms, and automation tooling. This is a senior-level, hands-on engineering role requiring a foundation in cloud infrastructure and reliability engineering, along with experience supporting edge technologies, networking, and operational automation. The ideal candidate can define reliability standards, measure system performance, and implement scalable solutions that continuously improve operational excellence. You will work with Reliability Engineering, DevOps, and Security teams while balancing strategic architecture responsibilities with hands-on engineering execution.



Job #

3054890



Job Description

Senior Site Reliability Engineer



Location

Louisville, Kentucky (Hybrid)



Key Responsibilities


  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets to balance reliability with velocity.

  • Design, build, and maintain monitoring, alerting, and observability platforms that provide meaningful signals.

  • Lead blameless post-incident reviews, drive root cause analysis, and ensure permanent corrective actions are implemented.

  • Design, implement, and support secure, scalable, and highly available cloud solutions primarily within AWS, with working knowledge of Azure.

  • Architect and manage containerized workloads using Kubernetes, including both cloud and edge-based deployments.

  • Develop and maintain Infrastructure as Code (IaC) using Terraform.

  • Build and optimize CI/CD pipelines using GitLab CI/CD and modern DevOps practices.

  • Develop automation solutions, internal tools, and middleware integrations to eliminate repetitive operational work.

  • Design and support Kubernetes-based edge systems deployed in restaurant locations.

  • Manage and optimize Mobile Device Management (MDM) platforms.

  • Establish and enforce cloud governance, security policies, and architectural standards.

  • Provide technical leadership and mentorship to engineering teams.



Required Qualifications


  • 6+ years in IT infrastructure, with 3+ years focused on site reliability engineering, cloud architecture, or platform engineering.

  • Hands-on experience with AWS services (e.g., VPC, EC2, ECS, EKS, Fargate, Lambda, IAM, RDS, S3).

  • Working experience with Microsoft Azure.

  • Strong expertise in Kubernetes and container orchestration, including edge or distributed deployments.

  • Experience with GitLab CI/CD and CI/CD pipeline design.

  • Solid experience with Terraform for infrastructure provisioning.

  • Experience with enterprise networking, including switch configuration, VLANs, and network troubleshooting.

  • Familiarity with Mobile Device Management (MDM) platforms.

  • Experience with automation and scripting (Python, Bash, Go, or equivalent).

  • Proven ability to build internal tooling and API integrations.

  • Experience defining and operating against SLOs, SLIs, and error budgets.

  • Comfortable working in Linux command-line environments.



Preferred Qualifications


  • Experience designing serverless architectures (AWS Lambda, Fargate, API Gateway, EventBridge).

  • Experience with DevSecOps tooling (SAST, DAST, container scanning, IaC scanning).

  • Familiarity with security frameworks (CIS, NIST, ISO 27001, SOC 2).

  • AWS and/or Azure certifications.

  • Experience with monitoring and observability tools (CloudWatch, Prometheus, Grafana, or SIEM solutions).

  • Background in restaurant, retail, or distributed edge technology environments.

  • Experience using AI-assisted development tools for scripting, troubleshooting, and documentation.



Key Competencies


  • Generalist mindset, comfortable moving between cloud, edge, networking, and device management.

  • Reliability-first thinking, treating toil reduction and error budgets as core engineering disciplines.

  • Bias toward permanent fixes and automation over repeated manual intervention.

  • Ability to balance strategic architecture with hands-on execution.

  • Strong communication and cross-functional collaboration skills.

  • A proactive and detail-oriented approach to systems design and operations.



Equal Opportunity Employer Statement

Everforth Apex Systems is an equal opportunity employer. We do not discriminate or allow discrimination on the basis of race, color, religion, creed, sex (including pregnancy, childbirth, breastfeeding, or related medical conditions), age, sexual orientation, gender identity, national origin, ancestry, citizenship, genetic information, registered domestic partner status, marital status, disability, status as a crime victim, protected veteran status, political affiliation, union membership, or any other characteristic protected by law. Everforth Apex will consider qualified applicants with criminal histories in a manner consistent with the requirements of applicable law.



Accommodation Statement

If you require an accommodation under the Americans with Disabilities Act to participate in an interview with a virtual recruiter or to use our website for a search or application, please contact our Benefits Department at [email protected] or 804-523-8228. Please note that this contact information is strictly to be used for medical ADA accommodations and that no other inquiries will be answered.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer-Developer 3
Software Engineer-Developer 3

Apex Systems • Colorado Springs (CO), Northern (KY)

On-site
USD 90,000 - 130,000
Medical insurance
Dental insurance
Vision insurance
+5
Infrastructure & Network Engineer
Infrastructure & Network Engineer

Apex Systems • Houston (TX)

On-site
USD 110,000 - 170,000
Medical
Dental
Vision
+7
Site Reliability Engineer Leader
Site Reliability Engineer Leader

Apex Systems • Hartford (CT)

On-site
USD 180,000 - 240,000
Medical, dental, vision
401K with company match
HSA and EAP
+1
Architect 1
Architect 1

Apex Systems • Raymond (OH)

On-site
USD 110,000 - 170,000
Medical, dental, vision
401K with company match
ESPP
+3
Chief Infrastructure Engineer
Chief Infrastructure Engineer

Apex Systems • Adelphi (MD)

On-site
USD 135,000 - 175,000
Medical insurance
Dental insurance
Vision insurance
+2
Cloud Full Stack Developer II
Cloud Full Stack Developer II

Apex Systems • Newport Beach (CA)

Hybrid
USD 146,000 - 154,000
Medical insurance
Dentist/vision coverage
Life and disability insurance
+2
Cloud Systems Engineer
Cloud Systems Engineer

Apex Systems • Pleasanton (CA)

Remote
USD 120,000 - 150,000
Medical coverage
Dental coverage
Vision coverage
+2
Site Operations Director (San Fran)
Site Operations Director (San Fran)

Apex Systems • San Francisco (CA)

On-site
USD 150,000 - 210,000
Health insurance
401K
ESPP
Software Engineer
Software Engineer

Apex Systems • Greenwood Village (CO)

Hybrid
USD 120,000 - 170,000
Frontend Developer III
Frontend Developer III

Apex Systems • Jersey City (NJ)

On-site
USD 140,000 - 180,000