Site Reliability Engineering (SRE) Architect

STAFFWORXS

Atlanta (GA)

Hybrid

USD 96,432 - 103,320

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A prominent recruitment firm is looking for a Site Reliability Engineering (SRE) Architect to lead the design and development of reliable and scalable systems in Atlanta, Georgia. This hybrid role focuses on defining architectural blueprints, enhancing system observability, and advocating best practices across engineering teams. Qualified candidates will have proven experience in reliability roles and a strong background in AWS, SRE principles, and system architecture. The position offers competitive pay based on skills and experience.

Qualifications

  • Proven experience in architectural roles focused on reliability and performance.
  • Deep hands-on expertise with SRE principles, including error budgets and incident management.
  • Strong AWS experience, especially in infrastructure and security.

Responsibilities

  • Architect scalable, highly available solutions on AWS.
  • Define SRE standards and evaluate current observability systems.
  • Serve as a senior advisor on scalability and performance across teams.

Skills

Architectural design for reliability
SRE principles expertise
AWS proficiency
Containerization (Kubernetes, Docker)
Observability solutions (Dynatrace, Prometheus)
Programming skills (Python, Go, Bash)
Strategic problem-solving
Leadership abilities

Job description

Site Reliability Engineering (SRE) Architect

Get AI-powered advice on this job and more exclusive features.

This range is provided by STAFFWORXS. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.

Base pay range

$70.00/hr - $75.00/hr

Direct message the job poster from STAFFWORXS

Delivery Manager @ STAFFWORXS | US IT Recruitment

Job Title: Site Reliability Engineering (SRE) Architect

Location: Atlanta, Georgia

Work Model: Hybrid (In-person presence required)

Overview

We are seeking a highly experienced Site Reliability Engineering (SRE) Architect to lead the strategic design, development, and maturity of our reliability engineering practices. This role goes beyond operational support, focusing on defining the architectural blueprint, standards, and frameworks that guide development and SRE operations teams in building resilient, scalable, and high-performing systems. The SRE Architect will influence technology decisions, enhance system observability, and foster a culture of reliability across the organization.

Key Responsibilities
  • Reliability Strategy & Architecture
    • Architect scalable, highly available, secure, and cost-effective solutions on AWS.
    • Define and promote SRE standards, best practices, and architectural blueprints across engineering teams.
    • Evaluate and enhance current observability systems, identifying gaps and driving next-level maturity to improve system insights.
    • Lead the definition and implementation of SLIs, SLOs, and error budgets for critical services.
    • Design solutions to eliminate operational toil through automation and improved system architecture.
    • Assess existing SRE tools, CI/CD pipelines, IaC modules, and automated remediation frameworks, proposing improvements.
    • Evaluate and recommend new tools, technologies, and practices to strengthen reliability, productivity, and operational excellence.
  • Technical Leadership & Consultation
    • Serve as a senior advisor on reliability, scalability, and performance across development and platform teams.
    • Offer architectural guidance for new services to ensure reliability principles are integrated from the start.
    • Mentor SREs and engineers, promoting strong engineering discipline and adherence to SRE principles.
    • Lead architecture reviews and production readiness assessments for critical systems.
  • Resilience Engineering
    • Lead blameless postmortems for major incidents and drive systemic architectural improvements.
    • Advocate and architect resilience patterns including circuit breakers, rate limiting, graceful degradation, and chaos engineering.
Required Qualifications
  • Proven experience in architectural roles focused on reliability, scalability, and performance.
  • Deep hands-on expertise with SRE principles (SLIs/SLOs, error budgets, automation, incident management).
  • Strong AWS experience across infrastructure, networking, and security.
  • Expertise with containerization and orchestration (Kubernetes, Docker, serverless).
  • Experience building observability solutions (Dynatrace, Prometheus, Grafana, ELK/EFK, Jaeger, OpenTelemetry).
  • Strong programming/scripting abilities (Python, Go, Bash).
  • Excellent analytical and strategic problem-solving skills.
  • Strong communication, collaboration, and leadership abilities.
Preferred Qualifications
  • Experience implementing and maturing chaos engineering practices and platforms.
Seniority level
  • Mid-Senior level
Employment type
  • Contract
Job function
  • Other

Referrals increase your chances of interviewing at STAFFWORXS by 2x

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Robotics Technologies LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Cloud Hybrid Technologies, LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Cloud Analytics Technologies, LLC • Atlanta (GA)

On-site
USD 140,000 - 210,000
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

MACHINE LEARNING TECHNOLOGIES LLC • Atlanta (GA)

On-site
USD 140,000 - 190,000
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Quantum Technologies. LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Robotics Prcocess Automation, LLC • Atlanta (GA)

On-site
USD 100,000 - 150,000
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Ethereum Technologies LLC • Atlanta (GA)

On-site
USD 83,000 - 165,000
Site Reliability Engineering (SRE) Consultant
Site Reliability Engineering (SRE) Consultant

TekWissen ® • Charlotte (NC)

On-site
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

Hybrid
USD 140,000 - 150,000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1
Site Reliability Engineer
Site Reliability Engineer

SRE • Puerto Rico

Hybrid
USD 120,000 - 180,000