Sr. Site Reliability Engineer (SRE)

Avenue Code

California (MO)

On-site

USD 150,000 - 200,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Avenue Code is seeking an experienced SRE to partner with product teams, design, build, and operate our cloud platform, and drive reliability, performance, and security across engineering. You will own infrastructure automation, incident response, and performance dashboards in a fast-paced consultancy environment.

Ideal candidates have 5+ years of production systems, strong AWS, Kubernetes, Terraform, and CI/CD experience, plus English proficiency.

Qualifications

  • 5+ years operating production systems.
  • Deep AWS Cloud and cloud-native practices.
  • Kubernetes (EKS/GKE) at scale experience.
  • Terraform for declarative provisioning.
  • Redis and PostgreSQL administration knowledge.
  • VPC, VPN, load balancers and cloud networking.
  • Git workflows, branching, and CI/CD integrations.
  • Strong HTTP/REST/TLS/DNS protocol knowledge.
  • Professional English proficiency.

Responsibilities

  • Automate provisioning and deployments with IaC and CI/CD (Terraform, GitHub Actions, ArgoCD).
  • Define SLIs/SLOs, manage error budgets, build dashboards and alerts.
  • Enforce IAM least-privilege policies and automate vulnerability scans.
  • Instrument services with metrics, logs, and distributed tracing.
  • Own on-call rotations, lead incident response and post-mortems.
  • Implement tagging and cost controls to optimize cloud spend.
  • Create runbooks and mentor teams on DevOps, reliability, security.

Education

Bachelor's degree in CS or related

Tools

GitHub Actions
ArgoCD
Jenkins
Terraform
Kubernetes
Redis
PostgreSQL
VPC networking
Git workflows

Job description

We’re seeking an experienced, highly collaborative SRE to partner with product teams and tackle our most critical infrastructure challenges. You’ll be hands‑on in designing, building, and operating our cloud platform—and driving the reliability, performance, and security that empower our engineering organization.

Responsibilities
  • Infrastructure as Code & CI/CD: Automate provisioning and deployments with
  • Terraform and integrate best-practice pipelines (GitHub Actions, ArgoCD, etc.).
  • Reliability Engineering: Define SLIs/SLOs, manage error budgets, and build dashboards & alerts to proactively measure and improve system health.
  • Security & Compliance: Enforce least‑privilege IAM policies, automate vulnerability scans, and maintain audit logging for compliance.
  • Monitoring & Observability: Instrument services with metrics, logs, and distributed tracing to enable rapid troubleshooting, aid teams in alerting, custom metrics, and dashboarding
  • Incident Management: Own on‑call rotations, lead real‑time incident response, conduct post‑mortems, and drive continuous improvements.
  • Cost Optimization: Implement tagging strategies, right‑size resources, and leverage concrete data to decide on optimal methods to control cloud spend at scale.
  • Documentation & Mentorship: Author runbooks, standards, and best‑practice guides—and coach dev teams on implementing modern DevOps, reliability, and security patterns.
Required Qualifications
  • Have 5+ years of experience running production critical systems.
  • Deep proficiency with the AWS Cloud and Cloud‑Native best practices.
  • Experience with Kubernetes (EKS, GKE) and Container Orchestration at scale.
  • Skilled in Terraform to declaratively provision and maintain infrastructure services.
  • Working knowledge of managing and debugging databases like Redis and Postgres.
  • Strong familiarity with VPC, VPN, Load Balancing, and cloud networking components.
  • Proficiency with Git workflows, branching strategies, and CI/CD systemintegrations.
  • Solid understanding of web and network protocols and standards (HTTP, REST, TLS, DNS, etc...)
  • Professional proficiency in English (both written and spoken) is required for this role.
Nice to Have Skills
  • Bachelor's degree, or equivalent in Computer Science, Engineering, or a related field.
  • Experience with ArgoCD, GitHub Actions, Jenkins, or other CI/CD pipeline solutions.
  • Working knowledge of Python, Golang, and Helm templating languages.
  • Node.js experience a plus, including running scalable, resilient Node microservices.
  • Grasp of foundational security best practices for cloud infrastructure.
  • Awareness of Terragrunt, managing Terraform state, and optimal project structure.
  • Seasoned in production readiness fundamentals amidst a fast-moving team.

Avenue Code discloses salary range information based on our commitment to fairness and transparency. We consider a wide range of factors such as internal equity, geographic location, relevant education, qualifications, certifications, experience, skills, seniority, business or organizational needs, and others. At Avenue Code, it is not typical for an individual to be hired at or near the top of the range for their role, and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range for a SRE POSITION is from $150.000,00 to $200.000,00 yearly.

Avenue Code reinforces its commitment to privacy and to all the principles guaranteed by the most accurate global data protection laws, such as GDPR, LGPD, CCPA and CPRA. The Candidate data shared with Avenue Code will be kept confidential and will not be transmitted to disinterested third parties, nor will it be used for purposes other than the application for open positions. As a Consultancy company, Avenue Code may share your information with its clients and other Companies from the CompassUol Group to which Avenue Code’s consultants are allocated to perform its services.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud SRE: Reliability, Security & Platform
Senior Cloud SRE: Reliability, Security & Platform

Avenue Code • California (MO)

On-site
USD 150,000 - 200,000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Weekday (YC W21) • New York (NY)

On-site
USD 150,000 - 250,000
Health, dental, vision insurance
Generous PTO
Learning & development
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Supio • San Francisco (CA)

On-site
USD 170,000 - 220,000
Site Reliability Engineer
Site Reliability Engineer

Stelvio Inc. • Town of Texas (WI)

On-site
USD 125,000 - 145,000
Site Reliability Engineer
Site Reliability Engineer

Axle • Frederick (MD)

On-site
USD 140,000 - 155,000
Paid Time Off
401K match
Educational Benefits
+5
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Randstad Digital Americas • Plano (TX)

On-site
USD 115,000 - 125,000
Medical insurance
401K plan
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Site Reliability Engineer
Site Reliability Engineer

VantageScore® • San Francisco (CA)

On-site
USD 150,000
Medical insurance
Dental insurance
401(k) plan
+1
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

Calabrio • United States

On-site
USD 90,000 - 130,000