Site Reliability Engineer

Staffworxs

Louisville (KY)

Hybrid

USD 120,000 - 180,000

Full time

6 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Staffworxs in Louisville, KY is seeking a Site Reliability Engineer with strong experience in AWS, Kubernetes and Terraform to own reliability, observability and cloud architectures for restaurant technology systems.

This 12-month contract role offers hybrid on-site work (on-site Tuesdays and Thursdays) with responsibilities spanning SRE, CI/CD, IaC, and security governance.

Qualifications

  • 6+ years in IT infrastructure with 3+ years in SRE/cloud/platf orm engineering.
  • Hands-on AWS experience (VPC, EC2, EKS, Fargate, IAM, RDS, S3).
  • Working experience with Microsoft Azure.
  • Strong Kubernetes expertise, including edge deployments.
  • Experience designing CI/CD pipelines with GitLab CI/CD.
  • Terraform for infrastructure provisioning.
  • Experience with enterprise networking and MDM.
  • Automation and scripting in Python/Bash/Go.

Responsibilities

  • Define and own SLOs/SLIs for restaurant tech; use error budgets.
  • Build and maintain monitoring, alerting, observability tools.
  • Lead blameless post-incident reviews and RCAs.
  • Identify reliability risks across the stack proactively.
  • Track reliability metrics; report system health to leadership.
  • Design scalable AWS cloud architectures; Azure knowledge helpful.
  • Architect/manage container workloads with Kubernetes (edge).
  • Develop serverless and container solutions (Fargate).
  • Maintain IaC using Terraform.
  • Own architecture across cloud, edge, networking, devices.

Skills

SRE experience
Cloud architecture
CI/CD
Automation
Python scripting
Edge deployments
Security basics
Linux CLI
Troubleshooting
Cross-functional collaboration

Tools

AWS
Azure
Kubernetes
Terraform
GitLab CI/CD
Terraform Cloud
Prometheus
Grafana
CloudWatch
MDM

Job description

At Staffworxs, we don't just connect talent — we power transformation. Headquartered in Frisco, TX, with teams in Bengaluru and Hyderabad, we combine global reach with deep expertise. Our Digital & Data Analytics practice drives growth and innovation for some of the world's top brands, who continue to retain us as their trusted partner. If you're ready to make an impact, you're in the right place.

Role :- Site Reliability Engineer with strong experience in AWS, Kubernetes and Terraform

Location :- Louisville, KY (Hybrid) (on site Tuesday and Thursday )

Duration: 12 Months Contract

Job Description
Key Responsibilities
Reliability & Observability
  • Define and own Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for restaurant technology systems; use error budgets to balance reliability with velocity.
  • Build and maintain monitoring, alerting, and observability platforms that provide meaningful signal — not noise.
  • Lead blameless post-incident reviews; drive root cause analysis and ensure permanent corrective actions are implemented.
  • Proactively identify reliability risks across the stack before they become incidents.
  • Establish and track reliability metrics; report on system health to engineering leadership.
  • Design and implement scalable, secure, and highly available cloud architectures primarily on AWS, with working knowledge of Azure.
  • Architect and manage containerized workloads using Kubernetes (EKS, AKS), including edge Kubernetes deployments in restaurant environments.
  • Design and implement serverless and container-based solutions, including AWS Fargate and other managed services.
  • Develop and maintain Infrastructure as Code (IaC) using Terraform.
  • Own architecture across the full restaurant technology stack — cloud, edge, networking, and device management — not just the cloud layer.
DevOps, Automation & Tooling
  • Build and optimize CI/CD pipelines using GitLab CI/CD and modern DevOps practices.
  • Build internal tools, automations, and middleware integrations that eliminate repetitive operational work.
  • Use AI-assisted development to accelerate scripting, troubleshooting, and documentation.
  • Champion a culture of engineering solutions over repeated manual fixes — if something is done twice, it should be automated.
  • Design and support Kubernetes-based edge systems deployed in restaurant locations.
  • Support mobile application deployments and troubleshoot deployment issues across restaurant endpoints.
  • Manage and optimize Mobile Device Management (MDM) platforms covering the restaurant device fleet.
  • Configure and troubleshoot enterprise networking — primarily switches and restaurant-facing network infrastructure.
  • Lead and participate in incident response for restaurant technology systems, including on-call coverage and post-incident review.
  • Reduce mean time to detection (MTTD) and mean time to resolution (MTTR) through better tooling, runbooks, and automation.
Security & Governance
  • Establish and enforce cloud governance, security policies, and architectural standards.
  • Implement cloud security best practices: IAM strategy, network segmentation, encryption, and secrets management.
  • Conduct security architecture reviews; identify vulnerabilities, misconfigurations, and compliance gaps.
  • Integrate security into CI/CD pipelines (DevSecOps — Development, Security, and Operations), including automated scanning, policy validation, and vulnerability management.
  • Collaborate with engineering, DevOps, and security teams to ensure secure-by-design solutions across cloud and restaurant tech.
  • Provide technical leadership and mentorship to engineering teams.
  • Create and maintain documentation, runbooks, and architectural decision records.
  • Continuously evaluate emerging technologies and recommend improvements.
Required Qualifications
  • 6+ years in IT infrastructure, with 3+ years focused on site reliability engineering, cloud architecture, or platform engineering.
  • Hands-on experience with AWS (VPC, EC2, ECS, EKS, Fargate, Lambda, IAM, RDS, S3).
  • Working experience with Microsoft Azure.
  • Strong expertise in Kubernetes and container orchestration, including edge or distributed deployments.
  • Experience with GitLab CI/CD and CI/CD pipeline design.
  • Solid experience with Terraform for infrastructure provisioning.
  • Experience with enterprise networking — switch configuration, VLANs, network troubleshooting.
  • Familiarity with Mobile Device Management (MDM) platforms.
  • Experience with automation and scripting (Python, Bash, Go, or equivalent).
  • Proven ability to build internal tooling and API integrations, not just configure managed services.
  • Experience defining and operating against SLOs, SLIs, and error budgets.
  • Comfortable working in Linux command-line environments; Windows familiarity a plus where restaurant endpoints require it.
Preferred Qualifications
  • Experience designing serverless architectures (AWS Lambda, Fargate, API Gateway, EventBridge).
  • Experience with DevSecOps tooling (SAST — Static Application Security Testing, DAST — Dynamic Application Security Testing, container scanning, IaC scanning).
  • Familiarity with security frameworks (CIS, NIST, ISO 27001, SOC 2).
  • AWS and/or Azure certifications.
  • Experience with monitoring and observability tools (CloudWatch, Prometheus, Grafana, or SIEM solutions).
  • Background in restaurant, retail, or distributed edge technology environments.
  • Experience using AI-assisted development tools for scripting, troubleshooting, and documentation.
  • Generalist mindset — comfortable moving between cloud, edge, networking, and device management in the same week.
  • Reliability-first thinking — treats toil reduction, error budgets, and post-incident learning as core engineering disciplines, not afterthoughts.
  • Bias toward permanent fixes and automation over repeated manual intervention.
  • Ability to balance strategic architecture with hands-on execution.
  • Strong communication and cross-functional collaboration skills.
  • Curious, proactive, and detail-oriented approach to systems design and operations.

Staffworxs is an Equal Opportunity Employer. We celebrate diversity and are committed to creating an inclusive workplace for all employees, regardless of race, color, religion, gender, sexual orientation, national origin, age, disability, or veteran status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineering (SRE)/Dev Ops
Site Reliability Engineering (SRE)/Dev Ops

Spectraforce • Louisville (KY)

Hybrid
USD 80,000 - 90,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Experis Technology Group • Louisville (KY)

Hybrid
USD 65,000 - 87,000
Medical and Prescription Drug Plans
Dental Plan
Vision Plan
+2
Site Reliability Engineering (SRE)/Dev Ops
Site Reliability Engineering (SRE)/Dev Ops

Pyramid Consulting, Inc • Louisville (KY)

On-site
USD 83,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

Spectraforce Technologies • Austin (TX)

On-site
USD 120,000 - 155,000
Site Reliability Engineer
Site Reliability Engineer

Pentangle Tech Services | P5 Group • Lowell (MA)

On-site
USD 140,000 - 190,000
Site Reliability Engineer II
Site Reliability Engineer II

Restaurant365 • Denver (CO)

On-site
USD 98,583 - 138,016
100% employee medical benefits
401k + matching
Equity Option Grant
+2
Site Reliability Engineer II
Site Reliability Engineer II

Restaurant365 • San Francisco (CA)

On-site
USD 98,583 - 138,016
Comprehensive medical benefits, 100% paid for employee
401k + matching
Equity Option Grant
+2
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Principal Reliability Engineer
Principal Reliability Engineer

GoTo Foods, LLC • Blythe (GA)

On-site
USD 180,000 - 240,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Clearwater Analytics • Boise (ID)

On-site
USD 130,000 - 170,000