Senior Site-Reliability Engineer

Jobgether SRL

United States

Remote

USD 105,000 - 140,000

Full time

3 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Remote work based in the United States
Health, dental, and vision insurance
401(k) retirement plan with company匹配
Paid time off and holidays

Job summary

Jobgether SRL in the United States is seeking a Senior Site-Reliability Engineer to ensure reliability, availability, and performance of Windows-based production environments. You will bridge development and operations, automate infrastructure, and optimize deployment pipelines in a hybrid setup.

The role emphasizes incident response, post-mortems, and continuous improvement with modern cloud, automation, and monitoring technologies.

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or related field preferred.
  • 5+ years in Systems Administration, DevOps, SRE or similar infrastructure role.
  • Strong Windows Server expertise (2016+), Active Directory, IIS, and MS SQL.

Responsibilities

  • Design, implement, and maintain scalable infrastructure using IaC across production.
  • Develop automation scripts in PowerShell, Python, Ruby, and more for provisioning and config management.
  • Manage configuration via Terraform, Puppet, and/or Chef in hybrid environments.
  • Monitor health, performance, reliability, and availability with observability tools.
  • Establish SLAs/SLOs and error budgets for production services.

Skills

Windows Server
AWS
Terraform
Puppet/Chef
Monitoring tools
CI/CD

Education

Bachelor's degree in CS/IT

Tools

Docker
Kubernetes
GitLab Pipelines
Jenkins
GitHub Actions

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site‑Reliability Engineer based in United States. This is a senior infrastructure engineering role focused on the reliability, availability, and performance of Windows‑based production environments. You will bridge development and operations to build highly available services and maintain strong operational standards across hybrid infrastructure. The role combines infrastructure automation, configuration management, observability, incident response, and continuous improvement. You will work closely with development teams to strengthen deployment and release processes while establishing measurable reliability objectives. The position offers an opportunity to work with modern cloud, automation, and monitoring technologies in a mission‑critical environment. The role is remote and requires the ability to obtain the appropriate security clearance.

Accountabilities
  • Design, implement, and maintain scalable infrastructure using Infrastructure as Code (IaC) practices across production environments.
  • Develop and maintain automation scripts using PowerShell, Python, Ruby, and other scripting languages for operating system provisioning, configuration management, and recurring operational tasks.
  • Implement and manage configuration management solutions such as Terraform, Puppet, and/or Chef across hybrid infrastructure environments.
  • Monitor system health, performance, reliability, and availability using established observability tools and operational best practices.
  • Establish, maintain, and enforce Service Level Agreements (SLAs), Service Level Objectives (SLOs), and error budgets for production services.
  • Participate in an on-call rotation, respond to production incidents, restore services rapidly, and conduct thorough root cause analysis.
  • Partner with development teams to improve deployment pipelines, release processes, and overall service reliability.
  • Create and maintain operational documentation, including procedures, runbooks, and architectural decisions.
  • Lead or contribute to post-mortem reviews and implement corrective actions designed to prevent recurring incidents.
  • Troubleshoot complex infrastructure and application issues across multiple technology layers while maintaining a strong focus on operational excellence.
Requirements
  • Bring at least 5 years of experience in Systems Administration, DevOps, Site‑Reliability Engineering, or a closely related infrastructure role.
  • Have strong hands‑on expertise with Windows Server environments, including Windows Server 2016 or later, Active Directory, IIS, and Microsoft SQL.
  • Demonstrate strong cloud infrastructure skills, with AWS experience preferred.
  • Possess advanced scripting capabilities, including the development of reusable modules and integrations with REST APIs.
  • Have hands‑on experience with Terraform for infrastructure provisioning and Puppet or Chef for configuration management.
  • Be experienced with monitoring and observability platforms such as Prometheus, Grafana, Datadog, or New Relic.
  • Have a solid understanding of networking fundamentals, including DNS, TCP/IP, load balancing, and VPN technologies.
  • Demonstrate strong analytical and problem‑solving skills, with the ability to troubleshoot complex issues spanning multiple technology layers.
  • A Bachelor's degree in Computer Science, Information Technology, or a related discipline is preferred, although equivalent professional experience may be considered.
  • Relevant certifications such as AWS Solutions Architect, Microsoft certifications, or HashiCorp Certified: Terraform Associate are desirable.
  • Experience with containerization technologies such as Docker and Kubernetes, as well as CI/CD tools including GitLab Pipelines, Jenkins, or GitHub Actions, is a plus.
  • Knowledge of security best practices, compliance frameworks, and log aggregation and analysis tools such as the ELK Stack or Splunk is desirable.
  • Be able to obtain the required security clearance for the position.
Benefits
  • Annual salary range of $105,000–$140,000, with actual compensation determined by experience, qualifications, skills, geographic location, contract requirements, and business needs.
  • Medical, dental, and vision insurance for eligible employees.
  • Life, AD&D, and disability insurance.
  • Paid time off and 11 company holidays.
  • 401(k) retirement plan with company matching.
  • Additional employee benefits and wellness resources, subject to applicable eligibility requirements and plan terms.
  • Remote work arrangement based in the United States.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

On-site
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Site Reliability Engineer
Site Reliability Engineer

Skill • Southlake (TX)

On-site
USD 66,000 - 73,000
Health insurance
Vision insurance
Dental insurance
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

On-site
USD 140,000 - 150,000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1
Senior Software Engineer, Site Reliability Engineering
Senior Software Engineer, Site Reliability Engineering

Ladders • United States

Remote
USD 180,000 - 233,000
Collaborative culture
Impact the platform's future direction
Engagement with diverse engineering
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

ConsultNet Technology Services and Solutions • El Segundo (CA)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer, Cloud Platform Infrastructure
Senior Site Reliability Engineer, Cloud Platform Infrastructure

Solü Technology Partners • Buffalo (NY)

On-site
USD 130,000 - 165,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GovCIO • Arlington (VA)

On-site
USD 210,000 - 230,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kontakt.io • New York (NY)

Hybrid
USD 200,000 - 240,000
Hybrid schedule
Equity
Health, dental, and vision
+3
Site Reliability Engineer
Site Reliability Engineer

Request Technology, LLC • Chicago (IL)

On-site
USD 150,000 - 155,000