Senior Site Reliability Engineer (SC Cleared)

Montash

United States

Remote

GBP 70,000 - 110,000

Full time

34 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Montash is seeking a Senior Site Reliability Engineer (SC Cleared) to work remotely on a government client project. You will drive SRE best practices across a large cloud estate, onboard services through assessment gates, and lead incident response with a strong focus on reliability and security.

The role requires hands-on engineering, coaching across teams, and participation in a 24/7 on-call rota, with influence on design, monitoring, and automation initiatives within production environments.

Qualifications

  • Extensive SRE experience driving reliability and observability.
  • Experience leading incident investigations and post-incident reviews.
  • Experience managing error budgets with product owners.
  • Ability to onboard services to production via stage gates.
  • Experience implementing monitoring and runbooks.
  • Ability to assess change impact with stakeholders.
  • Experience reducing toil through automation.
  • Strong communication and influence across teams.

Responsibilities

  • Drive adoption of SRE best practice across cloud estate and onboarding.
  • Design techniques to improve reliability, runbooks, and knowledge transfer.
  • Collaborate with development teams from design phase with guidance on best practice.
  • Promote engineering ownership and maintenance of live services.
  • Manage the error budget in alignment with the product owner.
  • Act as focal point for major/complex incidents and ensure resourcing.
  • Conduct reviews for high priority incidents and publish outcomes.
  • Assess change requests with stakeholders and authorize implementations.
  • Coach and mentor engineers in SRE practices.
  • Reduce toil and increase automation for reliability.
  • Capture stakeholder input and promote collaboration and innovation.
  • Provide on-call support and participate in out-of-hours rota.

Skills

SRE Experience
Incident Management
Observability
Automation
On-call Experience
Stakeholder Communication

Job description

Job Title: Senior Site Reliability Engineer (SC Cleared)

Location: Remote

Contract Length: 6 months (with scope to extend)

Start Date: November 2026

IR35: Inside

Interview Process: 1 Stage, MS Teams

Clearance Required: SC (active, unelapsed clearance required)

We are supporting a government client in hiring a Senior Site Reliability Engineer to join its SRE team, driving the adoption of SRE best practice across a large cloud estate.

Using both soft skills and technical experience, the successful candidates will work with application teams to ensure the client's standards and governance are met when onboarding services into the cloud, through a dedicated assessment stage gate process, ensuring applications satisfy the operational and security needs of running in production. Working with development teams from the design phase, they will help apply good practice and standards to application infrastructure, supporting reliable and secure solutions for the public.

This is a senior role that leads by example, providing technical direction and supporting other SREs within the team. It suits an engineer who can influence and coach development teams as well as respond hands-on to major incidents. Participation in an out-of-hours on-call rota is a requirement.

Key Responsibilities
  • Drive adoption of SRE best practice across the cloud estate, including onboarding services through an assessment stage gate process
  • Design and develop techniques for improving application reliability, runbooks, knowledge transfer across teams and ongoing SRE strategy within professional communities
  • Work collaboratively with development teams from the design phase, providing guidance on best practice and ensuring application monitoring is enabled
  • Push a mindset change within the organisation to foster engineering ownership, SRE best practice and the importance of the integrity and maintenance of the live service
  • Manage the error budget agreed with the product owner for the application, balancing work in alignment with it
  • Act as the focal point for the investigation and resolution of major or complex incidents, ensuring people with the right skills are proactively available to respond
  • Conduct reviews for all high priority and major incidents, ensuring they are completed quickly and published
  • Assess the impact of change requests in consultation with stakeholders, providing technical expertise and authorising the implementation of subsequent changes
  • Coach and mentor application development and operations engineers in the practice and techniques of SRE
  • Help reduce toil and increase automation, improving reliability and reducing time to live and spend on repetitive tasks
  • Routinely seek views and capture ideas from stakeholders and team members, encouraging collaboration and innovation
  • Provide on-call support to help restore services, and participate in an out-of-hours on-call rota
Essential Skills
  • Strong Site Reliability Engineering experience, driving the adoption of SRE best practice (reliability, observability, automation and operational ownership) across a cloud estate
  • Experience leading the investigation and resolution of major or complex incidents, including running post-incident reviews and publishing the outcomes quickly
  • Experience managing error budgets agreed with product owners, balancing reliability and delivery work accordingly
  • Experience working with development teams from the design phase to apply good practice, standards and governance to application infrastructure, including onboarding services to production through assessment or stage gate processes
  • Experience implementing monitoring and observability for applications, and producing runbooks and knowledge-transfer material
  • Experience assessing the impact of change requests in consultation with stakeholders and providing the technical expertise to authorise changes
  • Experience reducing toil and increasing automation to cut repetitive work and time to live
  • Experience coaching and mentoring development and operations engineers in SRE practice
  • Strong communication and stakeholder skills, with the ability to influence engineering culture and foster engineering ownership
  • Ability to provide on-call support and take part in an out-of-hours on-call rota
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE (SC Cleared) – Remote, Incident Leader
Senior SRE (SC Cleared) – Remote, Incident Leader

Montash • United States

Remote
GBP 70,000 - 110,000
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Good co India • United States

Remote
USD 120,000 - 160,000
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE) / Principal Site Reliability Engineer (SRE)

Mindlance • Irving (TX)

On-site
USD 120,000 - 160,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Prosum • Scottsdale (AZ)

On-site
USD 150,000 - 190,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG • Allen (TX)

On-site
USD 150,000 - 190,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

ecsfederal • Virginia (MN)

Hybrid
USD 118,000 - 177,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG • Raleigh (NC)

On-site
USD 140,000 - 190,000
Healthcare
Retirement planning
Volunteer days
+1
Site Reliability Engineer
Site Reliability Engineer

LA International • United States

Hybrid
USD 98,000 - 122,000
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000