A technology consulting firm in Atlanta is looking for an experienced Site Reliability Engineering (SRE) Architect. In this role, you will design and build scalable infrastructure, ensure system reliability, and define SRE best practices. Candidates should have proven architectural expertise, a strong understanding of SRE principles, and experience with cloud platforms like AWS. This position offers an hourly rate based on experience and emphasizes a culture of reliability and operational excellence.
Qualifications
Proven experience in an architectural role focused on reliability and scalability.
Deep understanding of SRE principles including error budgets and automation.
Expertise in AWS infrastructure and services.
Responsibilities
Architect and design scalable, secure infrastructure on AWS.
Define SRE best practices for the engineering organization.
Lead postmortems for significant incidents and prioritize improvements.
Skills
Reliability Strategy & Design
SRE principles
Cloud computing platforms
Containerization and orchestration
Observability solutions
Programming/Scripting (Python, Go, Bash)
Analytical and problem-solving skills
Strong communication and leadership skills
Tools
AWS
Kubernetes
Docker
Dynatrace
Grafana
Prometheus
ELK/EFK Stack
Jaeger
OpenTelemetry
Job description
A technology consulting firm in Atlanta is looking for an experienced Site Reliability Engineering (SRE) Architect. In this role, you will design and build scalable infrastructure, ensure system reliability, and define SRE best practices. Candidates should have proven architectural expertise, a strong understanding of SRE principles, and experience with cloud platforms like AWS. This position offers an hourly rate based on experience and emphasizes a culture of reliability and operational excellence.