Senior SRE - Remote, Unlimited PTO, Reliable Platforms

Ad Tech Industry

United States

Remote

USD 140,000 - 190,000

Full time

9 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Unlimited PTO
Hobby & team building budget allowance
Employee Support Program
Loss of family member financial aid
Employee Resource Groups

Job summary

Semrush is seeking a Senior Site Reliability Engineer to join the SRE Team, ensuring reliability of our scalable infrastructure and applications. You will identify potential failure points and implement solutions with cross-functional partners to boost uptime and performance.

The role involves building tooling in Go or Python, shaping SLOs, and mentoring engineers, with on-call rotations and collaboration across development teams to design resilient systems.

Qualifications

  • 3+ years of experience as a Site Reliability Engineer.
  • Experience with Kubernetes and Cloud providers.
  • Experience with engineering in Python or Go.
  • Strong understanding of what an application failure is and how to handle it.
  • Ability to debug applications using metrics.
  • Familiarity with traces, observability, and implementation quirks in code.
  • Willingness to be on call and work flexible hours.
  • Team player with good communication abilities.
  • GCP knowledge.

Responsibilities

  • Lead the changes in common engineering practices in the Company.
  • Induce application failures and work to recover them from that state.
  • Debug applications using metrics and add traces/metrics as needed.
  • Establish and refine SLOs, cost dashboards, and security hardening initiatives in partnership with stakeholders to guarantee service reliability and performance.
  • Collaborate with development teams to design and implement scalable, reliable, and efficient system architecture.
  • Designs full-stack platform solutions from concept to production.
  • Builds sophisticated tooling in Go/Python to automate operations.
  • Mentors engineers, interviews candidates, leads critical incidents.
  • On-call rotation: Typically one week every 2–3 weeks and may include overnight incidents.

Skills

Kubernetes
Cloud providers
Python
Go
Observability & tracing
On-call experience
Communication
GCP

Tools

Go
Python

Job description

Semrush is seeking a Senior Site Reliability Engineer to join the SRE Team, ensuring reliability of our scalable infrastructure and applications. You will identify potential failure points and implement solutions with cross-functional partners to boost uptime and performance.

The role involves building tooling in Go or Python, shaping SLOs, and mentoring engineers, with on-call rotations and collaboration across development teams to design resilient systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Platform Reliability & Incident Lead (Remote)
Senior SRE: Platform Reliability & Incident Lead (Remote)

Affirm, Inc. • Town of Poland (NY)

On-site
USD 32,000 - 48,000
Health insurance
Equity rewards
Flexible Spending Wallets
+1
Senior SRE – Remote, High-Impact Cloud Infra
Senior SRE – Remote, High-Impact Cloud Infra

Upserve • United States

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

Hybrid
USD 150,000 - 190,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

ACI Infotech • Seattle (WA), Northern (KY)

Hybrid
USD 120,000 - 170,000
Health insurance
Dental insurance
Vision insurance
+3
Senior SRE: Remote, High-Impact Cloud Reliability
Senior SRE: Remote, High-Impact Cloud Reliability

FullStack • Jefferson City (MO)

On-site
USD 100,000 - 150,000
Health, dental, and vision insurance
401(k) with 4% match
Paid Time Off
+1
Senior SRE - Remote, Multi-Cloud, High-Scale
Senior SRE - Remote, Multi-Cloud, High-Scale

Intuition Machines • United States

Remote
PHP 7,299,000 - 10,949,000
Fully remote position
Flexible working hours
Global team
+2
Senior SRE - Hybrid, Platform Reliability Lead
Senior SRE - Hybrid, Platform Reliability Lead

TransUnion LLC • Reston (VA)

Hybrid
USD 112,500 - 187,500
Day-one medical, dental, vision
Company-paid basic life/AD&D
12 weeks paid parental leave
+2
Senior SRE: Scale Reliability, Observability & CI/CD
Senior SRE: Scale Reliability, Observability & CI/CD

Breakout Tools • San Francisco (CA)

On-site
USD 120,000 - 160,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

OutSolve • Mission (KS)

Remote
USD 90,000 - 130,000
100% remote work environment
Competitive compensation
Professional development opportunities
+1
Senior SRE: Scale, Reliability & Observability Leader
Senior SRE: Scale, Reliability & Observability Leader

Brez Technology Inc. • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Private Medical, Dental and Vision Benefits
Retirement Savings plan with matching contributions
Workspace benefits for your home office
+4