Sr. Site Reliability Engineer

Socket.dev

Kentucky

Hybrid

USD 120,000 - 170,000

Full time

5 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Socket.dev is seeking a Sr. Site Reliability Engineer to join a small, high-impact infra team.

You will own critical projects from design through deployment, including a pivotal modernization and hosting migration initiative in your first year. The role emphasizes automation via Terraform, Ansible, and PowerShell DSC; CI/CD with GitHub Actions and Jenkins; observability with Prometheus, Grafana, and Datadog; applying SRE principles such as SLOs and error budgets.

Qualifications

  • 5+ years in SRE/DevOps with production Azure or AWS.
  • Strong Linux/Unix administration and networking troubleshooting.
  • Expertise in Terraform and CI/CD pipeline design.
  • Scripting in Python/Go/Bash or PowerShell.

Responsibilities

  • Lead modernization efforts and hosting migrations for hybrid infra.
  • Automate provisioning with IaC and optimize CI/CD pipelines.
  • Establish observability and incident response standards.
  • Apply SRE practices including SLOs and error budgets.
  • Collaborate with development and data teams to manage change.
  • Ensure security, backups, and disaster recovery strategies.

Skills

SRE/DevOps
Cloud engineering
Linux/Networking
Programming/Scripting

Tools

Terraform
Ansible
GitHub Actions
Jenkins
Docker
Kubernetes
Datadog
Prometheus
Grafana

Job description

Sr. Site Reliability Engineer

We are seeking an experienced Sr. Site Reliability Engineer to join a small, high-impact infrastructure team. This role blends software engineering and systems automation to scale reliable cloud and hybrid systems. You will own critical projects from design through deployment, specifically driving a pivotal infrastructure modernization andhosting migration initiative in your first year.

Responsibilities:
  • Infrastructure & Migration: Lead modernization efforts and hosting migrations while maintaining hybrid infrastructure (Azure, AWS, on-prem).
  • Automation & IaC: Streamline provisioning using Infrastructure as Code (Terraform, Ansible, PowerShell DSC) and enhance CI/CD pipelines(GitHub Actions, Jenkins)for rapid delivery.
  • Observability: Introduce and manage comprehensive monitoring platforms (e.g. Prometheus, Grafana, Datadog)to establish operational standards.
  • Reliability Engineering: Manage containerized work loads in Docker and implement SRE principles including SLOs and error budgets.
  • Operational Excellence: Lead incident response, conduct root cause analysis, and automate manual processes through scripting and runbooks.
  • Security & Continuity: Support DevSecOps initiatives and ensure robust backup and disaster recovery strategies.
  • Emerging Tech & R&D: Continuously evaluate and pilot emerging technologies, tools, and industry trends to ensure our infrastructure stack remains modern, efficient, and scalable.
Qualifications:
  • Minimum 5 years in SRE, DevOps, or Cloud Engineering with production experience in Azure or AWS.
  • Strong Linux/Unix administration and networking troubleshooting skills.
  • Expertise in Infrastructure as Code (Terraform) and CI/CD pipeline design.
  • Proficiency in scripting or programming (Python, Go, Bash, or PowerShell).
  • Expertise with Docker in production environments; Kubernetes experience is a strong plus.
  • Self-starter with strong communication and documentation skills; able to take ownership in a small-team environment.
  • Excellent oral and written communication skills. Able to communicate effectively with a diverse group of individuals with varying levels of technical understanding and varying skillsets.
  • Ability to work collaboratively with application development and data engineering areas to define standards and manage change.
  • Must reside in the Greater Cincinnati Metropolitan Area (Hybrid Schedule 3 days a week in office)
Preferred Qualifications:
  • Practical knowledge of observability tools (Prometheus, Grafana, ELK, or similar).
  • Windows Server/Active Directory administration.
  • Experience withlegacyUnix (AIX/Solaris).
  • Database (Oracle/MS-SQL) or BI platform experience (Snowflake/Azure Fabric).
  • Relevant industry certifications (Azure, AWS, or CKA).

Hybrid Schedule three days a week in office required

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mikealbert • Cincinnati (OH)

Hybrid
USD 100,000 - 130,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

On-site
USD 100,000 - 135,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior SRE - Azure
Senior SRE - Azure

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 100,000 - 140,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Eliassen Group • Town of Florida (NY)

Hybrid
Medical benefits
Dental benefits
Vision benefits
+2
Site Reliability Engineer
Site Reliability Engineer

Ethos Group • Irving (TX)

On-site
USD 110,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Sr SRE Automation Engineer
Sr SRE Automation Engineer

Compunnel, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 130,000 - 180,000
Senior SRE / DevOps Engineer
Senior SRE / DevOps Engineer

Compunnel, Inc. • Jersey City (NJ)

On-site
USD 100,000 - 130,000