Sr. Software Engineer – Site Reliability Engineering (SRE)
Location: Atlanta, GA | Hybrid
Details:
Hybrid (onsite Tues/Wed), Contract-to-Hire, 5+ years experience required, H1B not eligible.
We’re looking for a Senior Site Reliability Engineer (SRE) to join our team managing AWS infrastructure and deployment pipelines for multiple development teams.
What You’ll Do:
- Automate testing, deployment, monitoring, and recovery processes.
- Define SLOs, error budgets, and implement best practices with engineering teams.
- Design and maintain tools for reliable application delivery and performance.
- Improve predictability, reliability, and reduce Mean Time to Recovery (MTTR).
Required Skills:
- 5–7 years' experience in software development / architecture.
- Expertise with Terraform, AWS (ECS, Lambda), Docker, GitHub Actions, and Linux/Unix systems.
- Experience using AWS SDKs or APIs for infrastructure automation
- Experience with monitoring/observability tools (e.g., New Relic, Splunk, PagerDuty).
- Agile development, CI/CD, automated testing, and strong problem‑solving skills.
Preferred Skills:
- Broad AWS platform knowledge (Cognito, WAF, Elasticache, S3, SNS, SQS).
- Database experience (RDS, MySQL, Postgres).